readable AI benchmarks (simplified)
OpenAI
GPT-5.5 (Non-reasoning)
GPT-5.5 (Non-reasoning) has a Quality Score of 41.6, ranking 158th among 551 scored models. Per 1M tokens, pricing is $5.00 input, $30.00 output, and $0.500 cached input. It has a 922k token context window. Its Reliability Score is 47.6, ranking 139th among 553 scored models, above average (total average: 35.6). Its Value Score is 24.5, low compared with the total Value average of 30.7.
Quality Score41.6coverage 100.0%: AA Intelligence and AA-Omniscience outcomes available
Value Score24.5
Factual reliability47.6
Cache Discount90%
Model specification
- Reasoning
- non-reasoning
- Input modalities
- none
- Output modalities
- none
- Context window
- 922k tokens
- Weights
- closed weights
Token prices USD per 1M tokens
- Input
- $5.00
- Output
- $30.00
- Cached Input
- $0.500
- Cache write
- not available
- Cost per task
- not available
Capability
- Intelligence
- 23.2
- Coding
- not available
- Agentic
- not available
- Omniscience
- -4.8
- Correct
- 45.6%
- Blended price
- $4.35
Answer outcomes
Fully graded outcomes (Correct + Incorrect): 96.0%
Correct 45.6%Incorrect 50.4%Partial / not attempted 4.0%
Artificial Analysis benchmarks
- GDPval-AA v2
- 23%
- τ³-Banking
- 15%
- Humanity’s Last Exam
- 14%
- GPQA Diamond
- 77%
- CritPt
- 1%
- AA-Omniscience
- 48%
- AA-LCR
- 64%
Similar models
- GPT-6 Sol (Non-reasoning)OpenAIQuality Score 46.4Value Score 30.8
- GPT-5.6 Terra (Non-reasoning)OpenAIQuality Score 35.2Value Score 23.0
- GPT-5.6 Sol (Non-reasoning)OpenAIQuality Score 47.1Value Score 28.8
- GPT-5.4 (Non-reasoning)OpenAIQuality Score 35.1Value Score 22.3
- GPT-5.2 (Non-reasoning)OpenAIQuality Score 35.0Value Score 22.7
- GPT-6 Luna (Non-reasoning)OpenAIQuality Score 33.6Value Score 30.7