readable AI benchmarks (simplified)

readable AI benchmarks

Apodex

Apodex 1.1

reasoning · closed weights · Aug 30, 2026

Apodex 1.1 has a Quality Score of 44.7, ranking 62nd among 266 scored models. Per 1M tokens, pricing is $0.300 input, $3.00 output, and $0.030 cached input. It accepts text and image input, outputs text, and has a 256k token context window. Its Reliability Score is 39.0, ranking 99th among 266 scored models, above average (total average: 33.4). Its Value Score is 40.0, high compared with the total Value average of 27.7.

Quality Score44.7coverage 91.2%: missing Epoch General ECI, CursorBench 3.2 and DeepSWE v1.1
Value Score40.0
Reliability39.0
Cache Discount90%

Model specification

Reasoning
reasoning
Input modalities
text, image
Output modalities
text
Context window
256k tokens
Weights
closed weights

Token prices USD per 1M tokens

Input
$0.300
Output
$3.00
Cached Input
$0.030
Cached Output
not available
Cost per task
$0.231

Capability

Intelligence
44.0
Coding
60.8
Agentic
36.6
Omniscience
-21.9
Correct
31.7%
Blended price
$0.381

Answer outcomes

Attempt rate: 85.3%

Correct 31.7%Incorrect 53.6%Abstained 14.7%

Artificial Analysis benchmarks

GDPval-AA v2
42%
τ³-Banking
25%
Terminal-Bench v2.1
70%
SciCode
43%
Humanity’s Last Exam
34%
GPQA Diamond
86%
CritPt
5%
AA-Omniscience
39%
AA-LCR
75%

Similar models