readable AI benchmarks (simplified)

readable AI benchmarks

Z AI

GLM-4.6V (Non-reasoning)

non-reasoning · open weights · Dec 8, 2025

GLM-4.6V (Non-reasoning) has a Quality Score of 22.0, ranking 344th among 551 scored models. Per 1M tokens, pricing is $0.300 input, $0.900 output, and $0.178 cached input. It accepts text, image, and video input, outputs text, and has a 128k token context window. Its Reliability Score is 31.2, ranking 295th among 553 scored models, below average (total average: 35.6). Its Value Score is 17.9, low compared with the total Value average of 30.7.

Quality Score22.0coverage 100.0%: AA Intelligence and AA-Omniscience outcomes available
Value Score17.9
Factual reliability31.2
Cache Discount41%

Model specification

Reasoning
non-reasoning
Input modalities
text, image, video
Output modalities
text
Context window
128k tokens
Weights
open weights

Token prices USD per 1M tokens

Input
$0.300
Output
$0.900
Cached Input
$0.178
Cache write
not available
Cost per task
not available

Capability

Intelligence
8.4
Coding
not available
Agentic
not available
Omniscience
-37.6
Correct
17.4%
Blended price
$0.274

Answer outcomes

Fully graded outcomes (Correct + Incorrect): 72.4%

Correct 17.4%Incorrect 55.0%Partial / not attempted 27.6%

Artificial Analysis benchmarks

Humanity’s Last Exam
4%
GPQA Diamond
57%
CritPt
0%
AA-Omniscience
31%
AA-LCR
17%

Similar models

Support me! Patreon