readable AI benchmarks (simplified)

readable AI benchmarks

NVIDIA

Nemotron 3.5 Lightning

reasoning · open weights · Aug 11, 2026

Nemotron 3.5 Lightning has a Quality Score of 30.5, ranking 270th among 577 scored models. Per 1M tokens, pricing is $0.060 input, $0.200 output, and $0.050 cached input. It accepts text input, outputs text, and has a 1M token context window. Its Reliability Score is 41.1, ranking 234th among 579 scored models, above average (total average: 36.6). Its Value Score is 28.1, around average compared with the total Value average of 31.8.

Quality Score30.5coverage 100.0%: AA Intelligence and AA-Omniscience outcomes available
Value Score28.1
Factual reliability41.1
Cache Discount17%

Model specification

Reasoning
reasoning
Input modalities
text
Output modalities
text
Context window
1M tokens
Weights
open weights

Token prices USD per 1M tokens

Input
$0.060
Output
$0.200
Cached Input
$0.050
Cache write
not available
Cost per task
$0.093

Capability

Intelligence
12.9
Coding
not available
Agentic
not available
Omniscience
-17.7
Correct
14.4%
Blended price
$0.067

Answer outcomes

Fully graded outcomes (Correct + Incorrect): 46.6%

Correct 14.4%Incorrect 32.1%Partial / not attempted 53.4%

Artificial Analysis benchmarks

GDPval-AA v2
7%
τ³-Banking
9%
SciCode
32%
Humanity’s Last Exam
11%
GPQA Diamond
74%
CritPt
0%
AA-Omniscience
41%
AA-LCR
60%

Similar models

Support me! Patreon