readable AI benchmarks

Baidu

ERNIE 5.0 Thinking Preview

reasoning · closed weights · Nov 13, 2025

Quality Score20.7
Value Scorenot available
Reliability27.2
Cache Discountnot available

Model specification

Reasoning
reasoning
Input modalities
text, image, video
Output modalities
text
Context window
128k tokens
Weights
closed weights

Token prices USD per 1M tokens

Input
not available
Output
not available
Cached Input
not available
Cached Output
not available
Cost per task
not available

Capability

Intelligence
22.3
Coding
not available
Agentic
not available
Omniscience
-45.6
Correct
22.1%
Blended price
not available

Answer outcomes

Correct 22.1%Incorrect 67.8%Abstained 10.1%

Artificial Analysis benchmarks

SciCode
38%
Humanity’s Last Exam
13%
GPQA Diamond
78%
CritPt
1%
AA-Omniscience
27%
AA-LCR
7%

Similar models