Z.ai · GLM-5
Z.ai's text-only flagship: the same base as GLM-5.2 with post-training only, much stronger at long-horizon coding and cyber tasks, and thinking always on (low/high/max effort).
Leader: Claude Opus 5.5 at 58
Leader: Claude Mythos 5.1 at 60.9%
Leader: Claude Fable 5.1 at 91.4%
Leader: Claude Opus 5.5 at 66.9%
Leader: Claude Opus 5.5 at 1846
Leader: GPT-6 Astra at 96.1%
Leader: Claude Opus 5.5 at 61.4%
Leader: GPT-5.6 Sol at 32.3%
Leader: Kimi K3 at 88.7%
Leader: Claude Fable 5 at 1506
Leader: Claude Opus 5.5 at 1818
| Result | Score | Reported by | Source |
|---|---|---|---|
| Terminal-Bench 3.0 · Claude Code harness, avg@3 | 28.3% | vendor | GLM-5.3 model card (Hugging Face) ↗ |
| DeepSWE v1.1 · mini-swe-agent harness | 66.9% | vendor | GLM-5.3 model card (Hugging Face) ↗ |
| Agents' Last Exam (ALE-CLI) | 28.5% | vendor | GLM-5.3 model card (Hugging Face) ↗ |
| CyberGym | 84.5% | vendor | GLM-5.3 model card (Hugging Face) ↗ |
| Toolathlon Verified | 73% | vendor | GLM-5.3 model card (Hugging Face) ↗ |
| AutomationBench v1.0.6 | 48.2% | vendor | GLM-5.3 model card (Hugging Face) ↗ |
| NL2Repo | 58 score | vendor | GLM-5.3 model card (Hugging Face) ↗ |