readable AI benchmarks (simplified)

readable AI benchmarks

Anthropic

Claude 4 Opus (Non-reasoning)

non-reasoning · closed weights · May 22, 2025

Claude 4 Opus (Non-reasoning) does not yet have enough benchmark data for a Quality Score. Per 1M tokens, pricing is $15.00 input, $75.00 output, $1.50 cached input, and $18.75 cache write. It has a 200k token context window.

Quality Scorenot availableQuality unavailable: missing AA-Omniscience outcomes
Value Scorenot available
Factual reliabilitynot available
Cache Discount90%

Model specification

Reasoning
non-reasoning
Input modalities
none
Output modalities
none
Context window
200k tokens
Weights
closed weights

Token prices USD per 1M tokens

Input
$15.00
Output
$75.00
Cached Input
$1.50
Cache write
$18.75
Cost per task
not available

Capability

Intelligence
16.6
Coding
not available
Agentic
not available
Omniscience
not available
Correct
not available
Blended price
$11.55

Answer outcomes

Fully graded outcomes (Correct + Incorrect): not available

Omniscience outcome data not available for this model.

Artificial Analysis benchmarks

Humanity’s Last Exam
6%
GPQA Diamond
70%

Similar models

Support me! Patreon