modelbenchmark.io

Ling 3.0 Flash VL

InclusionAI · ling-3-0-flash-vl

Compare

Ling 3.0 Flash VL is inclusionAI's native multimodal Mixture-of-Experts model with 124B total parameters and 5.5B active parameters per token. It combines image and video understanding with reasoning and tool use for document analysis, charts, visual verification, and interface-based agent tasks. Thinking is enabled by default and can be turned off in settings.

Specification

most-agreed values

Context
262K
Max output
33K
Released
2026-09-09
Knowledge cutoff
Retires
Open weights
yes
Input
text, image, video
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.06
Output
$0.18
Cache read
$0.012

Available from 6 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
kilo · freefreefree262K33K
openrouter · freefreefree262K33K
vercel · inclusionaifreefree256K32K
kilo · inclusionai$0.06$0.18$0.012131K33K
nano-gpt · inclusionai$0.06$0.18$0.012262K33K
openrouter · inclusionai$0.06$0.18$0.012131K33K

More from InclusionAI

most-hosted first