9 ms·
True. Z.AI ran that bench themselves and report 46.2, which is lower than GPT-5.5 and Opus 4.8, but crushing the other open weights models. https://z.ai/blog/g
by ttul 4mo ago
True. Z.AI ran that bench themselves and report 46.2, which is lower than GPT-5.5 and Opus 4.8, but crushing the other open weights models.
https://z.ai/blog/glm-5.2 https://z.ai/blog/glm-5.2