Qwen omni model for text, vision, audio, and multimodal agent tasks
Specification
most-agreed values
Context
66K
Max output
16K
Released
2025-09-24
Knowledge cutoff
—
Retires
—
Open weights
yes
Input
text, audio, video, image
Output
text
Price
US dollars per million tokens · most-agreed
Input
$0.25
Output
$0.97
Available from 2 hosts
More from Alibaba
most-hosted first