readable AI benchmarks

Google

Gemini 3.5 Flash (medium)

reasoning · closed weights · May 19, 2026

Quality Score60.3
Value Score41.6
Reliability60.4
Cache Discount90%

Model specification

Reasoning
reasoning
Input modalities
text, image, video, speech
Output modalities
text
Context window
1M tokens
Weights
closed weights

Token prices USD per 1M tokens

Input
$1.50
Output
$9.00
Cached Input
$0.150
Cached Output
not available
Cost per task
not available

Capability

Intelligence
46.7
Coding
not available
Agentic
not available
Omniscience
20.8
Correct
51.0%
Blended price
$1.31

Answer outcomes

Correct 51.0%Incorrect 30.2%Abstained 18.7%

Artificial Analysis benchmarks

SciCode
53%
Humanity’s Last Exam
41%
GPQA Diamond
92%
CritPt
11%
AA-Omniscience
60%
AA-LCR
80%

Similar models