readable AI benchmarks (simplified)
Alibaba
Qwen3.8-Flash-Next
Qwen3.8-Flash-Next has a Quality Score of 53.7, ranking 43rd among 264 scored models. It accepts text, image, and video input, outputs text, and has a 256k token context window. Its Reliability Score is 45.1, ranking 67th among 264 scored models, above average (total average: 33.3).
Quality Score53.7coverage 91.2%: missing Epoch General ECI, CursorBench 3.2 and DeepSWE v1.1
Value Scorenot available
Reliability45.1
Cache Discountnot available
Model specification
- Reasoning
- reasoning
- Input modalities
- text, image, video
- Output modalities
- text
- Context window
- 256k tokens
- Weights
- open weights
Token prices USD per 1M tokens
- Input
- not available
- Output
- not available
- Cached Input
- not available
- Cached Output
- not available
- Cost per task
- not available
Capability
- Intelligence
- 55.8
- Coding
- 73.1
- Agentic
- 56.4
- Omniscience
- -9.7
- Correct
- 24.5%
- Blended price
- not available
Answer outcomes
Attempt rate: 58.7%
Correct 24.5%Incorrect 34.2%Abstained 41.3%
Artificial Analysis benchmarks
- GDPval-AA v2
- 62%
- τ³-Banking
- 45%
- Terminal-Bench v2.1
- 86%
- SciCode
- 47%
- Humanity’s Last Exam
- 38%
- GPQA Diamond
- 92%
- CritPt
- 11%
- AA-Omniscience
- 45%
- AA-LCR
- 77%
Similar models
- Qwen3.8 MaxAlibabaQuality Score 59.9Value Score 44.6
- Qwen3.8 2.4T A95BAlibabaQuality Score 60.9Value Score 44.3
- Qwen3.8 27B (xhigh)AlibabaQuality Score 47.2Value Score 41.5
- Qwen3.7 PlusAlibabaQuality Score 43.5Value Score 36.5
- Qwen3.8 27B (medium)AlibabaQuality Score 38.6Value Score 36.3
- Qwen3.8 27B (low)AlibabaQuality Score 38.5Value Score 35.1