3 ms·
Great work! Do you know if it's possible to port this over to pytorch's Apple Silicon/MPS support?
by smith7018 3y ago
Great work! Do you know if it's possible to port this over to pytorch's Apple Silicon/MPS support?
- chillee 3y agoUnfortunately it's a little bit tricky today. The main issue is that we rely heavily on torch.compile + Triton for performance in this repo, and there isn't an Apple Silicon backend either for torch.compile or Triton. For example, there's an AMD backend for Triton (and it's also integrated into torch.compile), which is why we can mostly do the same optimizations on Nvidia and AMD GPUs. Ideally, there'd be an Apple Silicon backend for Triton, and then this repo would mostly work out of the box :)