readable AI benchmarks

Alibaba

Qwen3 Next 80B A3B (Reasoning)

reasoning · open weights · Sep 11, 2025

Quality Score15.5
Value Score17.5
Reliability24.3
Cache Discountnot available

Model specification

Reasoning
reasoning
Input modalities
text
Output modalities
text
Context window
262,144 tokens
Weights
open weights

Token prices USD per 1M tokens

Input
$0.150
Output
$1.20
Cached Input
not available
Cached Output
not available
Cost per task
$0.042

Capability

Intelligence
16.9
Coding
17.4
Agentic
2.1
Omniscience
-51.4
Correct
19.0%
Blended price
$0.255

Answer outcomes

Correct 19.0%Incorrect 70.5%Abstained 10.5%

Artificial Analysis benchmarks

GDPval-AA v2
0%
τ³-Banking
6%
Terminal-Bench v2.1
7%
SciCode
39%
Humanity’s Last Exam
13%
GPQA Diamond
76%
CritPt
0%
AA-Omniscience
24%
AA-LCR
62%

Similar models