readable AI benchmarks (simplified)

readable AI benchmarks

Kimi

Kimi K3 (Low)

reasoning · open weights · release date not listed

Kimi K3 (Low) has a Quality Score of 49.1, ranking 110th among 571 scored models. Per 1M tokens, pricing is $3.00 input, $15.00 output, and $0.300 cached input. It has a 1,048,576 token context window. Its Reliability Score is 52.0, ranking 106th among 573 scored models, above average (total average: 36.5). Its Value Score is 31.1, around average compared with the total Value average of 31.4.

Quality Score49.1coverage 100.0%: AA Intelligence and AA-Omniscience outcomes available
Value Score31.1
Factual reliability52.0
Cache Discount90%

Model specification

Reasoning
reasoning
Input modalities
none
Output modalities
none
Context window
1,048,576 tokens
Weights
open weights

Token prices USD per 1M tokens

Input
$3.00
Output
$15.00
Cached Input
$0.300
Cache write
not available
Cost per task
$1.15

Capability

Intelligence
30.1
Coding
not available
Agentic
not available
Omniscience
3.9
Correct
45.8%
Blended price
$2.31

Answer outcomes

Fully graded outcomes (Correct + Incorrect): 87.6%

Correct 45.8%Incorrect 41.8%Partial / not attempted 12.4%

Artificial Analysis benchmarks

GDPval-AA v2
31%
τ³-Banking
42%
SciCode
53%
Humanity’s Last Exam
25%
GPQA Diamond
84%
CritPt
3%
AA-Omniscience
52%
AA-LCR
79%

Similar models

Support me! Patreon