readable AI benchmarks (simplified)

readable AI benchmarks

InclusionAI

Ling 3.0 Flash

reasoning · open weights · Aug 4, 2026

Ling 3.0 Flash has a Quality Score of 36.0, ranking 224th among 571 scored models. Per 1M tokens, pricing is $0.075 input, $0.220 output, and $0.015 cached input. It has a 262,144 token context window. Its Reliability Score is 41.1, ranking 230th among 573 scored models, above average (total average: 36.5). Its Value Score is 33.9, around average compared with the total Value average of 31.4.

Quality Score36.0coverage 100.0%: AA Intelligence and AA-Omniscience outcomes available
Value Score33.9
Factual reliability41.1
Cache Discount80%

Model specification

Reasoning
reasoning
Input modalities
none
Output modalities
none
Context window
262,144 tokens
Weights
open weights

Token prices USD per 1M tokens

Input
$0.075
Output
$0.220
Cached Input
$0.015
Cache write
not available
Cost per task
not available

Capability

Intelligence
20.1
Coding
not available
Agentic
not available
Omniscience
-17.9
Correct
18.2%
Blended price
$0.048

Answer outcomes

Fully graded outcomes (Correct + Incorrect): 54.2%

Correct 18.2%Incorrect 36.0%Partial / not attempted 45.8%

Artificial Analysis benchmarks

GDPval-AA v2
22%
τ³-Banking
27%
SciCode
42%
Humanity’s Last Exam
24%
GPQA Diamond
85%
CritPt
2%
AA-Omniscience
41%
AA-LCR
73%

Similar models

Support me! Patreon