3 ms·
It would appear that Flash-3 is already something that exists for PyTorch based on this joint blog between Nvidia, Together.ai and Princeton about enabling Flas
by sunshinesfbay 2y ago
It would appear that Flash-3 is already something that exists for PyTorch based on this joint blog between Nvidia, Together.ai and Princeton about enabling Flash-3 for PyTorch: https://pytorch.org/blog/flashattention-3/ https://pytorch.org/blog/flashattention-3/
- JackYoustra 2y agoRight - my point about "follows the same path" mostly revolves around llama.cpp's latency in adopting it.