The LLM benchmark aggregator.
GLM-5.3
Z.ai · 753B · open weights
Strongest in Agentic coding (#8 of 157, on 3 of its 4 boards), weakest in Mathematics (#86 of 140). Above par in 23 of 26 scopes. Among the models it meets almost everywhere, it finishes behind Claude Fable 5.1 and Muse Spark 1.3 and ahead of Claude Opus 4.5 and Claude Sonnet 5.
§ 1 · Profile
What it is good at
Bars run from 50 — the average of the models each source lists — so right of the line is above par. The middle column is the gap to whoever leads that scope.
§ 2 · Head to head
What it beats, and what beats it
The same models turn up scope after scope. Counted once: where both were placed, who finished higher. Bars run right for GLM-5.3, left for the other.
§ 3 · Sources
Where the numbers come from
9 publications, 39 figures. Every one links to the page it was read from.
Badge
[](https://publicai.io/model-index/m/glm-5-3)