readable AI benchmarks

Alibaba

Qwen3.5 397B A17B (Reasoning)

reasoning · open weights · Feb 16, 2026

Quality Score35.0
Value Score29.3
Reliability34.6
Cache Discountnot available

Model specification

Reasoning
reasoning
Input modalities
text, image
Output modalities
text
Context window
262,144 tokens
Weights
open weights

Token prices USD per 1M tokens

Input
$0.600
Output
$3.60
Cached Input
not available
Cached Output
not available
Cost per task
$0.357

Capability

Intelligence
34.3
Coding
48.2
Agentic
19.8
Omniscience
-30.8
Correct
30.8%
Blended price
$0.900

Answer outcomes

Correct 30.8%Incorrect 61.5%Abstained 7.7%

Artificial Analysis benchmarks

GDPval-AA v2
23%
τ³-Banking
13%
Terminal-Bench v2.1
51%
SciCode
42%
Humanity’s Last Exam
29%
GPQA Diamond
89%
CritPt
2%
AA-Omniscience
35%
AA-LCR
73%

Similar models