GLM-5 has 744B parameters and scores worse on MMLU-Pro than a 9B model
Summary
GLM-5, a 744B parameter model, underperforms on the MMLU-Pro benchmark compared to a much smaller 9B model, raising questions about efficiency and scaling.
Similar Articles
@j_golebiowski: A 1.7B parameter model beats GLM-5 (744B) on Schema Guided Dialogue — even when the training data is corrupted. That's …
A 1.7B parameter model surpasses 744B GLM-5 on Schema Guided Dialogue despite corrupted training data, showing 437× size efficiency.
@AdinaYakup: GLM 5.2 is here 753B ( smaller than you expect? ) 1M context MIT license GLM IndexShare: reuses the indexer across laye…
GLM 5.2 is released as a 753B parameter open-source model with 1M context length, MIT license, and achieves 99.2 on AIME 2026, outperforming GPT-5.5, Gemini 3.1 Pro, and Claude Opus 4.8.
GLM-5.2 is the new leading open weights model on Artificial Analysis
Z ai's GLM-5.2 has become the new leading open weights model on the Artificial Analysis Intelligence Index, scoring 51 and outperforming competitors like MiniMax-M3 and DeepSeek V4 Pro. The model features 744B total parameters, 40B active, MIT license, and 1M context window.
Human Evaluation of GLM-5.2
The author praises GLM-5.2, an MIT open-weights model, for its exceptional real-world performance in human evaluation benchmarks, claiming it rivals the best closed-source models like those from Claude.
Effect of GLM 5.2 !!
GLM 5.2, a new version of the GLM language model, has been released, demonstrating improved performance.