readable AI benchmarks (simplified)

readable AI benchmarks

MiniMax

MiniMax-M3

reasoning · open weights · Jun 1, 2026

MiniMax-M3 has a Quality Score of 47.8, ranking 123rd among 572 scored models. Per 1M tokens, pricing is $0.300 input, $1.20 output, $0.060 cached input, and $0.375 cache write. It accepts text, image, and video input, outputs text, and has a 1M token context window. Its Reliability Score is 50.7, ranking 117th among 574 scored models, above average (total average: 36.5). Its Value Score is 39.9, high compared with the total Value average of 31.5.

Quality Score47.8coverage 100.0%: AA Intelligence and AA-Omniscience outcomes available
Value Score39.9
Factual reliability50.7
Cache Discount80%

Model specification

Reasoning
reasoning
Input modalities
text, image, video
Output modalities
text
Context window
1M tokens
Weights
open weights

Token prices USD per 1M tokens

Input
$0.300
Output
$1.20
Cached Input
$0.060
Cache write
$0.375
Cost per task
$0.508

Capability

Intelligence
29.2
Coding
not available
Agentic
not available
Omniscience
1.4
Correct
16.7%
Blended price
$0.222

Answer outcomes

Fully graded outcomes (Correct + Incorrect): 32.0%

Correct 16.7%Incorrect 15.3%Partial / not attempted 68.0%

Artificial Analysis benchmarks

GDPval-AA v2
37%
τ³-Banking
15%
SciCode
47%
Humanity’s Last Exam
39%
GPQA Diamond
93%
CritPt
4%
AA-Omniscience
51%
AA-LCR
83%

Similar models

Support me! Patreon