3 ms·
How can it be that a 438B model is worse than Qwen3.8-27B? Are these benchmarks totally gamed?
by amazingamazing 24d ago
How can it be that a 438B model is worse than Qwen3.8-27B? Are these benchmarks totally gamed?
- magicalhippo 24d agoTraining data plays a huge role. As an example, Qwen 3 was generally considered a significant improvement over Qwen 2.5, but the architecture only had minor tweaks. The major change was the quantity and quality of training data they used for Qwen 3.
- ThouYS 24d agoqwen is a magical model. it has the mandate of heaven. was so already at 3.6-27B
- Caius-Cosades 24d ago[dead]