readable AI benchmarks (simplified)
Z AI
GLM-4.5 (Reasoning)
GLM-4.5 (Reasoning) has a Quality Score of 28.0, ranking 271st among 551 scored models. It accepts text input, outputs text, and has a 128k token context window. Its Reliability Score is 36.3, ranking 249th among 553 scored models, above average (total average: 35.6).
Quality Score28.0coverage 100.0%: AA Intelligence and AA-Omniscience outcomes available
Value Scorenot available
Factual reliability36.3
Cache Discountnot available
Model specification
- Reasoning
- reasoning
- Input modalities
- text
- Output modalities
- text
- Context window
- 128k tokens
- Weights
- open weights
Token prices USD per 1M tokens
- Input
- not available
- Output
- not available
- Cached Input
- not available
- Cache write
- not available
- Cost per task
- not available
Capability
- Intelligence
- 12.8
- Coding
- not available
- Agentic
- not available
- Omniscience
- -27.4
- Correct
- 25.1%
- Blended price
- not available
Answer outcomes
Fully graded outcomes (Correct + Incorrect): 77.6%
Correct 25.1%Incorrect 52.5%Partial / not attempted 22.3%
Artificial Analysis benchmarks
- Humanity’s Last Exam
- 13%
- GPQA Diamond
- 78%
- CritPt
- 0%
- AA-Omniscience
- 36%
- AA-LCR
- 53%
Similar models
- GLM-4.6V (Reasoning)Z AIQuality Score 26.9Value Score 21.9
- GLM-4.6 (Reasoning)Z AIQuality Score 28.8Value Score 22.4
- GLM-4.7-Flash (Reasoning)Z AIQuality Score 20.8Value Score 19.3
- GLM-4.7 (Reasoning)Z AIQuality Score 33.0Value Score 24.3
- GLM-4.5-AirZ AIQuality Score 18.1Value Score not available
- GLM-4.5V (Reasoning)Z AIQuality Score 19.2Value Score 15.1