4 ms·
Its not a new model, but rather their infrastructure and hardware they are showcasing.
by TechDebtDevin 1y ago
Its not a new model, but rather their infrastructure and hardware they are showcasing.
- pr337h4m 1y agoGroq appears to have quantized the Kimi K2 model they're serving, which is part of the reason why there's a noticeable performance gap between K2 on Moonshot's official API and the one served by Groq. We don't know how/whether the Qwen3-235B served by Cerebras has been quantized.
- logicchains 1y agoCerebras have previously stated for other models they hosted that they didn't quantise, unlike Groq.