readable AI benchmarks (simplified)
OpenAI
o3-pro
o3-pro does not yet have enough benchmark data for a Quality Score. Per 1M tokens, pricing is $20.00 input and $80.00 output. It accepts text and image input, outputs text, and has a 200k token context window.
Quality Scorenot availableQuality unavailable: missing AA-Omniscience outcomes
Value Scorenot available
Factual reliabilitynot available
Cache Discountnot available
Model specification
- Reasoning
- reasoning
- Input modalities
- text, image
- Output modalities
- text
- Context window
- 200k tokens
- Weights
- closed weights
Token prices USD per 1M tokens
- Input
- $20.00
- Output
- $80.00
- Cached Input
- not available
- Cache write
- not available
- Cost per task
- not available
Capability
- Intelligence
- 21.9
- Coding
- not available
- Agentic
- not available
- Omniscience
- not available
- Correct
- not available
- Blended price
- not available
Answer outcomes
Fully graded outcomes (Correct + Incorrect): not available
Omniscience outcome data not available for this model.
Artificial Analysis benchmarks
- GPQA Diamond
- 85%
Similar models
- GPT-5.5 Instant (May 2026)OpenAIQuality Score 45.2Value Score 26.6
- GPT-5.6 Luna (low)OpenAIQuality Score 37.5Value Score 32.0
- GPT-6 Luna (low)OpenAIQuality Score 38.8Value Score 35.5
- GPT-5 (medium)OpenAIQuality Score 39.9Value Score 26.9
- GPT-5 (low)OpenAIQuality Score 38.3Value Score 25.9
- GPT-5 (high)OpenAIQuality Score 40.5Value Score 27.3