readable AI benchmarks (simplified)

readable AI benchmarks

SpaceXAI

Grok 3 Reasoning Beta

reasoning · closed weights · Feb 19, 2025

Grok 3 Reasoning Beta does not yet have enough benchmark data for a Quality Score. It has a 1M token context window.

Quality Scorenot availableQuality unavailable: missing AA-Omniscience outcomes
Value Scorenot available
Factual reliabilitynot available
Cache Discountnot available

Model specification

Reasoning
reasoning
Input modalities
none
Output modalities
none
Context window
1M tokens
Weights
closed weights

Token prices USD per 1M tokens

Input
not available
Output
not available
Cached Input
not available
Cache write
not available
Cost per task
not available

Capability

Intelligence
10.4
Coding
not available
Agentic
not available
Omniscience
not available
Correct
not available
Blended price
not available

Answer outcomes

Fully graded outcomes (Correct + Incorrect): not available

Omniscience outcome data not available for this model.

Similar models

Support me! Patreon