3 ms·
Diffusion is already being used in Drafter in many LLMs. many people are running Qwen 3.8 27b on TPU at 130tk/s for free on Kaggle TPUs: https://www.reddit.co
by faangguyindia 26d ago
Diffusion is already being used in Drafter in many LLMs.
many people are running Qwen 3.8 27b on TPU at 130tk/s for free on Kaggle TPUs:
https://www.reddit.com/r/Qwen_AI/comments/1w6gv32/qwen3827b_at_130_toks_with_full_262k_context_on/ https://www.reddit.com/r/Qwen_AI/comments/1w6gv32/qwen3827b_...
I wonder if we are going to see boxes appear soon, which can run these models for dirt cheap.