Ling 3.0 Flash VL
InclusionAI · ling-3-0-flash-vl
Ling 3.0 Flash VL is inclusionAI's native multimodal Mixture-of-Experts model with 124B total parameters and 5.5B active parameters per token. It combines image and video understanding with reasoning and tool use for document analysis, charts, visual verification, and interface-based agent tasks. Thinking is enabled by default and can be turned off in settings.
Specification
most-agreed values
Context
262K
Max output
33K
Released
2026-09-09
Knowledge cutoff
—
Retires
—
Open weights
yes
Input
text, image, video
Output
text
Price
US dollars per million tokens · most-agreed
Input
$0.06
Output
$0.18
Cache read
$0.012
Available from 6 hosts
| Host | In $/M | Out $/M | Cache rd | Cache wr | Context | Output | Retires |
|---|---|---|---|---|---|---|---|
| kilo · free | free | free | — | — | 262K | 33K | — |
| openrouter · free | free | free | — | — | 262K | 33K | — |
| vercel · inclusionai | free | free | — | — | 256K | 32K | — |
| kilo · inclusionai | $0.06 | $0.18 | $0.012 | — | 131K | 33K | — |
| nano-gpt · inclusionai | $0.06 | $0.18 | $0.012 | — | 262K | 33K | — |
| openrouter · inclusionai | $0.06 | $0.18 | $0.012 | — | 131K | 33K | — |
More from InclusionAI
most-hosted first