Open-weight experimental preview of the Qwen4 architecture: hybrid-attention MoE (125B total, 6B active) with vision encoder for coding, agent tasks, and image and video understanding
Specification
most-agreed values
Context
262K
Max output
262K
Released
2026-08-27
Knowledge cutoff
—
Retires
—
Open weights
yes
Input
text, image, video
Output
text
Price
US dollars per million tokens · most-agreed
Input
$0.20
Output
$0.50
Cache read
$0.05
Quality
1 benchmark · 1 source
Composite
76th
percentile of 334 scored models
Rank
81 / 334
±0.53 sd
Evidence
1 × 1
too little evidence to place with confidence
Effort range
—
one configuration only
general61st1/1 bench
Every score
one row per source and configuration — nothing averaged away
| Benchmark | Score | Configuration | Source | Run |
|---|---|---|---|---|
| LiveBench | 77.3 | default · 23 runs | LiveBench | 2026-06-25 |
The leading figure for a benchmark is the median across its configurations, so one heroic high-effort run cannot set the number.
Available from 3 hosts
More from Alibaba
most-hosted first