readable AI benchmarks

Alibaba

Qwen3.5 397B A17B (Non-reasoning)

non-reasoning · open weights · Feb 16, 2026

Quality Score27.7
Value Score26.8
Reliability31.1
Cache Discountnot available

Model specification

Reasoning
non-reasoning
Input modalities
text, image
Output modalities
text
Context window
262,144 tokens
Weights
open weights

Token prices USD per 1M tokens

Input
$0.600
Output
$3.60
Cached Input
not available
Cached Output
not available
Cost per task
not available

Capability

Intelligence
32.7
Coding
not available
Agentic
not available
Omniscience
-37.9
Correct
24.5%
Blended price
$0.900

Answer outcomes

Correct 24.5%Incorrect 62.4%Abstained 13.1%

Artificial Analysis benchmarks

SciCode
41%
Humanity’s Last Exam
20%
GPQA Diamond
86%
CritPt
1%
AA-Omniscience
31%
AA-LCR
62%

Similar models