Pick this when
Cheap GLM coding when 5.3 is more than the job needs.
Skip when
You need the 5.3 score or a permissive license.
The downloadable GLM. Not the flagship. Useful if you already standardized on the family.
Price
- Input / 1M
- $0.30
- Output / 1M
- $1.00
- Measured task
- $0.19
Open-weight flash tier. Promo pricing reported through early September 2026.
Shape
- Context
- 1M
- Max output
- 64K
- Released
- 2026-08-18
- Tools
- Yes
- Reasoning
- Yes
Scores
Intelligence index52
GPQA Diamond81
Terminal-Bench 2.167
Agentic index39
- Arena
- —
- Speed
- 140/s
- GPQA
- 81.4%
- Term.
- 66.8%
- AA Index. The single number most people mean by “how smart.” It is a blend, so a specialist can lose here and still win the job you care about.
- GPQA. A clean test of hard reasoning. PhD experts sit around 65%. It says little about writing, tools, or taste.
- Term.. Closest public proxy for coding agents that live in a terminal. A high index score with a weak terminal score is a warning.
- Arena. Captures taste and usefulness that unit tests miss. Sample size varies by model, so treat gaps under ~100 Elo as noise.
Z.AI · Beijing. GLM. 5.3 matches Kimi on the composite at a lower API price. Weights are not public, so this is still a rental.