Qwen omni model for text, vision, audio, and multimodal agent tasks
Specification
most-agreed values
Context
66K
Max output
16K
Released
2025-09-24
Knowledge cutoff
2024-04
Retires
—
Open weights
yes
Input
text, video, audio, image
Output
text, audio
Price
US dollars per million tokens · most-agreed
Input
$0.25
Output
$0.97
Available from 2 hosts
More from Alibaba
most-hosted first