readable AI benchmarks

Alibaba

Qwen3.5 122B A10B (Non-reasoning)

non-reasoning · open weights · Feb 24, 2026

Quality Score24.2
Value Score23.2
Reliability22.5
Cache Discountnot available

Model specification

Reasoning
non-reasoning
Input modalities
text, image
Output modalities
text
Context window
262,144 tokens
Weights
open weights

Token prices USD per 1M tokens

Input
$0.400
Output
$3.20
Cached Input
not available
Cached Output
not available
Cost per task
$0.195

Capability

Intelligence
28.2
Coding
43.3
Agentic
16.2
Omniscience
-55.0
Correct
19.1%
Blended price
$0.680

Answer outcomes

Correct 19.1%Incorrect 74.1%Abstained 6.8%

Artificial Analysis benchmarks

GDPval-AA v2
19%
τ³-Banking
10%
Terminal-Bench v2.1
47%
SciCode
36%
Humanity’s Last Exam
16%
GPQA Diamond
83%
CritPt
1%
AA-Omniscience
23%
AA-LCR
62%

Similar models