Low-latency Gemini model for high-volume multimodal and agent workloads
Specification
most-agreed values
Context
1M
Max output
66K
Released
2026-05-07
Knowledge cutoff
2025-01
Retires
—
Open weights
no
Input
text, image, video, audio, pdf
Output
text
Price
US dollars per million tokens · most-agreed
Input
$0.25
Output
$1.50
Cache read
$0.025
Cache write
$0.0833
Available from 1 host
| Host | In $/M | Out $/M | Cache rd | Cache wr | Context | Output | Retires |
|---|---|---|---|---|---|---|---|
| edenai | $0.25 | $1.50 | $0.025 | $0.0833 | 1M | 66K | — |
More from Google
most-hosted first