readable AI benchmarks (simplified)

readable AI benchmarks

Meta

Muse Glimmer (High)

reasoning · open weights · Aug 10, 2026

Muse Glimmer (High) has a Quality Score of 30.2, ranking 265th among 571 scored models. Per 1M tokens, pricing is $0.325 input, $1.35 output, and $0.040 cached input. It has a 131,072 token context window. Its Reliability Score is 33.6, ranking 289th among 573 scored models, below average (total average: 36.5). Its Value Score is 25.1, low compared with the total Value average of 31.4.

Quality Score30.2coverage 100.0%: AA Intelligence and AA-Omniscience outcomes available
Value Score25.1
Factual reliability33.6
Cache Discount88%

Model specification

Reasoning
reasoning
Input modalities
none
Output modalities
none
Context window
131,072 tokens
Weights
open weights

Token prices USD per 1M tokens

Input
$0.325
Output
$1.35
Cached Input
$0.040
Cache write
not available
Cost per task
$0.057

Capability

Intelligence
17.5
Coding
not available
Agentic
not available
Omniscience
-32.9
Correct
27.0%
Blended price
$0.228

Answer outcomes

Fully graded outcomes (Correct + Incorrect): 86.8%

Correct 27.0%Incorrect 59.8%Partial / not attempted 13.2%

Artificial Analysis benchmarks

GDPval-AA v2
15%
τ³-Banking
24%
SciCode
45%
Humanity’s Last Exam
22%
GPQA Diamond
84%
CritPt
3%
AA-Omniscience
34%
AA-LCR
83%

Similar models

Support me! Patreon