readable AI benchmarks (simplified)
Anthropic
Claude Sonnet 5 (Non-reasoning, High)
Claude Sonnet 5 (Non-reasoning, High) has a Quality Score of 42.7, ranking 165th among 572 scored models. Per 1M tokens, pricing is $2.00 input, $10.00 output, $0.200 cached input, and $2.50 cache write. It has a 1M token context window. Its Reliability Score is 49.7, ranking 132nd among 574 scored models, above average (total average: 36.5). Its Value Score is 28.3, around average compared with the total Value average of 31.5.
Quality Score42.7coverage 100.0%: AA Intelligence and AA-Omniscience outcomes available
Value Score28.3
Factual reliability49.7
Cache Discount90%
Model specification
- Reasoning
- non-reasoning
- Input modalities
- none
- Output modalities
- none
- Context window
- 1M tokens
- Weights
- closed weights
Token prices USD per 1M tokens
- Input
- $2.00
- Output
- $10.00
- Cached Input
- $0.200
- Cache write
- $2.50
- Cost per task
- not available
Capability
- Intelligence
- 23.2
- Coding
- not available
- Agentic
- not available
- Omniscience
- -0.7
- Correct
- 33.8%
- Blended price
- $1.54
Answer outcomes
Fully graded outcomes (Correct + Incorrect): 68.3%
Correct 33.8%Incorrect 34.5%Partial / not attempted 31.7%
Artificial Analysis benchmarks
- GDPval-AA v2
- 36%
- τ³-Banking
- 16%
- Humanity’s Last Exam
- 19%
- GPQA Diamond
- 80%
- CritPt
- 1%
- AA-Omniscience
- 50%
- AA-LCR
- 70%
Similar models
- Claude Sonnet 4.6 (Non-reasoning, Low)AnthropicQuality Score 42.4Value Score 26.8
- Claude Opus 4.5 (Non-reasoning)AnthropicQuality Score 42.2Value Score 25.2
- Claude Sonnet 4.6 (Non-reasoning, High)AnthropicQuality Score 43.1Value Score 27.3
- Claude Opus 4.6 (Non-reasoning, High)AnthropicQuality Score 45.9Value Score 27.4
- Claude 4.5 Sonnet (Non-reasoning)AnthropicQuality Score 37.5Value Score 23.8
- Claude 4 Sonnet (Non-reasoning)AnthropicQuality Score 35.5Value Score not available