readable AI benchmarks (simplified)
Z AI
GLM-5.1 (Non-reasoning)
GLM-5.1 (Non-reasoning) has a Quality Score of 38.0, ranking 190th among 551 scored models. Per 1M tokens, pricing is $1.38 input, $4.40 output, and $0.260 cached input. It has a 200k token context window. Its Reliability Score is 38.8, ranking 228th among 553 scored models, above average (total average: 35.6). Its Value Score is 27.0, around average compared with the total Value average of 30.7.
Quality Score38.0coverage 100.0%: AA Intelligence and AA-Omniscience outcomes available
Value Score27.0
Factual reliability38.8
Cache Discount81%
Model specification
- Reasoning
- non-reasoning
- Input modalities
- none
- Output modalities
- none
- Context window
- 200k tokens
- Weights
- open weights
Token prices USD per 1M tokens
- Input
- $1.38
- Output
- $4.40
- Cached Input
- $0.260
- Cache write
- not available
- Cost per task
- not available
Capability
- Intelligence
- 24.2
- Coding
- not available
- Agentic
- not available
- Omniscience
- -22.4
- Correct
- 25.2%
- Blended price
- $0.898
Answer outcomes
Fully graded outcomes (Correct + Incorrect): 72.8%
Correct 25.2%Incorrect 47.6%Partial / not attempted 27.2%
Artificial Analysis benchmarks
- Humanity’s Last Exam
- 28%
- GPQA Diamond
- 84%
- CritPt
- 0%
- AA-Omniscience
- 39%
- AA-LCR
- 53%
Similar models
- GLM-5 (Non-reasoning)Z AIQuality Score 38.8Value Score 28.6
- GLM-5.2 (Non-reasoning)Z AIQuality Score 40.6Value Score 28.8
- GLM-4.6 (Non-reasoning)Z AIQuality Score 28.6Value Score 21.3
- GLM-4.7 (Non-reasoning)Z AIQuality Score 26.5Value Score 20.0
- GLM 5V Turbo (Reasoning)Z AIQuality Score 38.3Value Score not available
- GLM-5-TurboZ AIQuality Score 41.4Value Score not available