readable AI benchmarks

NVIDIA

Llama 3.1 Nemotron Instruct 70B

non-reasoning · open weights · Oct 15, 2024

Quality Score13.2
Value Score8.6
Reliability29.6
Cache Discountnot available

Model specification

Reasoning
non-reasoning
Input modalities
text
Output modalities
text
Context window
128k tokens
Weights
open weights

Token prices USD per 1M tokens

Input
$1.20
Output
$1.20
Cached Input
not available
Cached Output
not available
Cost per task
not available

Capability

Intelligence
7.4
Coding
not available
Agentic
not available
Omniscience
-40.8
Correct
17.8%
Blended price
$1.20

Answer outcomes

Correct 17.8%Incorrect 58.6%Abstained 23.6%

Artificial Analysis benchmarks

SciCode
23%
Humanity’s Last Exam
4%
GPQA Diamond
46%
CritPt
0%
AA-Omniscience
30%
AA-LCR
7%

Similar models