readable AI benchmarks (simplified)
Z AI
Celeris
Celeris has a Quality Score of 16.3, ranking 458th among 551 scored models. Per 1M tokens, pricing is $0.600 input, $1.80 output, and $0.110 cached input. It has a 64k token context window. Its Reliability Score is 22.2, ranking 428th among 553 scored models, below average (total average: 35.6). Its Value Score is 12.8, low compared with the total Value average of 30.7.
Quality Score16.3coverage 100.0%: AA Intelligence and AA-Omniscience outcomes available
Value Score12.8
Factual reliability22.2
Cache Discount82%
Model specification
- Reasoning
- non-reasoning
- Input modalities
- none
- Output modalities
- none
- Context window
- 64k tokens
- Weights
- open weights
Token prices USD per 1M tokens
- Input
- $0.600
- Output
- $1.80
- Cached Input
- $0.110
- Cache write
- not available
- Cost per task
- not available
Capability
- Intelligence
- 6.7
- Coding
- not available
- Agentic
- not available
- Omniscience
- -55.6
- Correct
- 18.2%
- Blended price
- $0.377
Answer outcomes
Fully graded outcomes (Correct + Incorrect): 92.0%
Correct 18.2%Incorrect 73.8%Partial / not attempted 8.0%
Artificial Analysis benchmarks
- Humanity’s Last Exam
- 3%
- GPQA Diamond
- 57%
- CritPt
- 0%
- AA-Omniscience
- 22%
- AA-LCR
- 0%
Similar models
- GLM-4.7-Flash (Non-reasoning)Z AIQuality Score 15.9Value Score 14.5
- GLM-4.6V (Non-reasoning)Z AIQuality Score 22.0Value Score 17.9
- GLM-4.7 (Non-reasoning)Z AIQuality Score 26.5Value Score 20.0
- GLM-4.6 (Non-reasoning)Z AIQuality Score 28.6Value Score 21.3
- GLM-4.5V (Reasoning)Z AIQuality Score 19.2Value Score 15.1
- GLM-4.5-AirZ AIQuality Score 18.1Value Score not available