modelbenchmark.io

Qwen Flash

Alibaba · qwen-flash

Compare

Efficient Qwen model for fast chat, extraction, and high-volume workloads

Specification

alibaba listing

Context
1M
Max output
33K
Released
2025-07-28
Knowledge cutoff
2024-04
Retires
Open weights
no
Input
text
Output
text

Price

US dollars per million tokens · alibaba list price

Input
$0.05
Output
$0.40
Cache read
$0.01
Cache write
$0.0625

Available from 11 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
alibabavendor$0.05$0.401M33K
alibaba-cn$0.022$0.2161M33K
merge-gateway$0.022$0.216$0.00441M250K
ofox$0.022$0.22$0.0043$0.0271M32K
ofox · bailian$0.022$0.22$0.0043$0.0271M33K
llmgateway$0.05$0.40$0.01$0.06251M33K
llmgateway-providers$0.05$0.40$0.01$0.06251M32K
llmtr$0.05$0.401M33K
dashscope · dashscope998K33K
qwen_ai_platform · qwen-ai-platform998K33K
qwencloud · qwencloud998K33K

More from Alibaba

most-hosted first