readable AI benchmarks

Alibaba

Qwen3.8 27B (low)

reasoning · open weights · Aug 14, 2026

Quality Score38.5
Value Score35.1
Reliability36.7
Cache Discount90%

Model specification

Reasoning
reasoning
Input modalities
text, image, video
Output modalities
text
Context window
256k tokens
Weights
open weights

Token prices USD per 1M tokens

Input
$0.500
Output
$3.00
Cached Input
$0.050
Cached Output
not available
Cost per task
$0.166

Capability

Intelligence
42.9
Coding
58.2
Agentic
43.7
Omniscience
-26.7
Correct
17.1%
Blended price
$0.435

Answer outcomes

Correct 17.1%Incorrect 43.8%Abstained 39.1%

Artificial Analysis benchmarks

GDPval-AA v2
49%
τ³-Banking
32%
Terminal-Bench v2.1
67%
SciCode
40%
Humanity’s Last Exam
14%
GPQA Diamond
85%
CritPt
0%
AA-Omniscience
37%
AA-LCR
75%

Similar models