GLM-5 has 744B parameters and scores worse on MMLU-Pro than a 9B model

Reddit r/artificial Models

Summary

GLM-5, a 744B parameter model, underperforms on the MMLU-Pro benchmark compared to a much smaller 9B model, raising questions about efficiency and scaling.

No content available
Original Article

Similar Articles

GLM-5.2 is the new leading open weights model on Artificial Analysis

Hacker News Top

Z ai's GLM-5.2 has become the new leading open weights model on the Artificial Analysis Intelligence Index, scoring 51 and outperforming competitors like MiniMax-M3 and DeepSeek V4 Pro. The model features 744B total parameters, 40B active, MIT license, and 1M context window.

Human Evaluation of GLM-5.2

Reddit r/LocalLLaMA

The author praises GLM-5.2, an MIT open-weights model, for its exceptional real-world performance in human evaluation benchmarks, claiming it rivals the best closed-source models like those from Claude.

Effect of GLM 5.2 !!

Reddit r/singularity

GLM 5.2, a new version of the GLM language model, has been released, demonstrating improved performance.