modelbenchmark.io

Qwen3.8 Flash Next (EU)

Alibaba · qwen3-8-flash-next-eu

Compare

Open-weight experimental preview of the Qwen4 architecture: hybrid-attention MoE (125B total, 6B active) with vision encoder for coding, agent tasks, and image and video understanding

Specification

most-agreed values

Context
262K
Max output
262K
Released
2026-08-27
Knowledge cutoff
Retires
Open weights
yes
Input
text, image, video
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.20
Output
$0.50
Cache read
$0.05

Available from 1 host

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
requesty$0.20$0.50$0.05262K262K

More from Alibaba

most-hosted first