readable AI benchmarks (simplified)

readable AI benchmarks

DeepSeek

DeepSeek V4 Pro 0424 (Non-reasoning)

non-reasoning · open weights · release date not listed

DeepSeek V4 Pro 0424 (Non-reasoning) has a Quality Score of 33.6, ranking 243rd among 571 scored models. Per 1M tokens, pricing is $0.435 input, $0.870 output, and $0.0036 cached input. It has a 1M token context window. Its Reliability Score is 35.1, ranking 278th among 573 scored models, below average (total average: 36.5). Its Value Score is 28.6, around average compared with the total Value average of 31.4.

Quality Score33.6coverage 100.0%: AA Intelligence and AA-Omniscience outcomes available
Value Score28.6
Factual reliability35.1
Cache Discount99%

Model specification

Reasoning
non-reasoning
Input modalities
none
Output modalities
none
Context window
1M tokens
Weights
open weights

Token prices USD per 1M tokens

Input
$0.435
Output
$0.870
Cached Input
$0.0036
Cache write
not available
Cost per task
not available

Capability

Intelligence
20.8
Coding
not available
Agentic
not available
Omniscience
-29.9
Correct
30.9%
Blended price
$0.177

Answer outcomes

Fully graded outcomes (Correct + Incorrect): 91.6%

Correct 30.9%Incorrect 60.7%Partial / not attempted 8.4%

Artificial Analysis benchmarks

Humanity’s Last Exam
8%
GPQA Diamond
72%
CritPt
1%
AA-Omniscience
35%
AA-LCR
53%

Similar models

Support me! Patreon