readable AI benchmarks (simplified)
Google
Gemini 2.5 Pro
Gemini 2.5 Pro has a Quality Score of 33.3, ranking 227th among 551 scored models. Per 1M tokens, pricing is $1.25 input, $10.00 output, and $0.125 cached input. It accepts text, image, audio, video, and pdf input, outputs text, and has a 1M token context window. Its Reliability Score is 41.8, ranking 198th among 553 scored models, above average (total average: 35.6). Its Value Score is 22.5, low compared with the total Value average of 30.7.
Quality Score33.3coverage 100.0%: AA Intelligence and AA-Omniscience outcomes available
Value Score22.5
Factual reliability41.8
Cache Discount90%
Model specification
- Reasoning
- reasoning
- Input modalities
- text, image, audio, video, pdf
- Output modalities
- text
- Context window
- 1M tokens
- Weights
- closed weights
Token prices USD per 1M tokens
- Input
- $1.25
- Output
- $10.00
- Cached Input
- $0.125
- Cache write
- not available
- Cost per task
- $0.232
Capability
- Intelligence
- 16.1
- Coding
- not available
- Agentic
- not available
- Omniscience
- -16.4
- Correct
- 39.1%
- Blended price
- $1.34
Answer outcomes
Fully graded outcomes (Correct + Incorrect): 94.5%
Correct 39.1%Incorrect 55.4%Partial / not attempted 5.5%
Artificial Analysis benchmarks
- GDPval-AA v2
- 0%
- τ³-Banking
- 10%
- SciCode
- 46%
- Humanity’s Last Exam
- 23%
- GPQA Diamond
- 84%
- CritPt
- 3%
- AA-Omniscience
- 42%
- AA-LCR
- 69%
Similar models
- Gemini 3.1 Flash-LiteGoogleQuality Score 32.9Value Score 27.5
- Gemini 2.5 Flash Preview (Sep '25) (Reasoning)GoogleQuality Score 27.9Value Score not available
- Gemma 4 31B (Reasoning)GoogleQuality Score 27.6Value Score not available
- Gemini 2.5 Flash (Reasoning)GoogleQuality Score 27.6Value Score 22.1
- Gemma 4 26B A4B (Reasoning)GoogleQuality Score 25.1Value Score 22.5
- Gemma 4 E4B (Reasoning)GoogleQuality Score 26.9Value Score not available