readable AI benchmarks (simplified)
Anthropic
Claude Haiku 5.5 (Max, Default Fallback)
Claude Haiku 5.5 (Max, Default Fallback) has a Quality Score of 61.0, ranking 60th among 577 scored models. Per 1M tokens, pricing is $0.100 input, $0.500 output, $0.010 cached input, and $0.125 cache write. It accepts text, image, and pdf input, outputs text, and has a 1M token context window. Its Reliability Score is 55.3, ranking 91st among 579 scored models, above average (total average: 36.6). Its Value Score is 55.8, high compared with the total Value average of 31.8.
Quality Score61.0coverage 100.0%: AA Intelligence and AA-Omniscience outcomes available
Value Score55.8
Factual reliability55.3
Cache Discount90%
Model specification
- Reasoning
- reasoning
- Input modalities
- text, image, pdf
- Output modalities
- text
- Context window
- 1M tokens
- Weights
- closed weights
Token prices USD per 1M tokens
- Input
- $0.100
- Output
- $0.500
- Cached Input
- $0.010
- Cache write
- $0.125
- Cost per task
- $0.213
Capability
- Intelligence
- 43.4
- Coding
- not available
- Agentic
- not available
- Omniscience
- 10.7
- Correct
- 36.4%
- Blended price
- $0.077
Answer outcomes
Fully graded outcomes (Correct + Incorrect): 62.1%
Correct 36.4%Incorrect 25.7%Partial / not attempted 37.9%
Artificial Analysis benchmarks
- GDPval-AA v2
- 56%
- SciCode
- 55%
- Humanity’s Last Exam
- 44%
- CritPt
- 19%
- AA-Omniscience
- 55%
- AA-LCR
- 83%
Similar models
- Claude Sonnet 5.5 (Medium, Default Fallback)AnthropicQuality Score 61.4Value Score 40.8
- Claude Opus 4.7 (Max)AnthropicQuality Score 63.1Value Score 37.7
- Claude Opus 5 (Low)AnthropicQuality Score 62.4Value Score 37.2
- Claude Haiku 5.5 (Xhigh, Default Fallback)AnthropicQuality Score 58.2Value Score 53.2
- Claude Opus 4.8 (Max)AnthropicQuality Score 64.3Value Score 38.4
- Claude Sonnet 5 (Max)AnthropicQuality Score 58.5Value Score 38.8