2 ms·
Aw very interesting! This is great feedback - what model are you running? I'm keen to do more crowdsourced data as time goes on.
by rlindsey123 18d ago
Aw very interesting! This is great feedback - what model are you running? I'm keen to do more crowdsourced data as time goes on.
- shadowpho 18d agoThe big three :) Qwen3.8-flash-next Deepseek4-0731-flash Glm5.3 The latest unsloth llama.cpp has a lot of nice features that runs them faster than before. I’ll have to double check which one runs how fast, but it’s generally 20-40 t/s. (And infil is fast but not sure how that’s counted)