modelbenchmark.io

Gemini 3.1 Flash Lite (Vertex AI, US)

Google · gemini-3-1-flash-lite-us

Compare

Low-latency Gemini model for high-volume multimodal and agent workloads

Specification

most-agreed values

Context
1M
Max output
66K
Released
2026-05-07
Knowledge cutoff
2025-01
Retires
Open weights
no
Input
text, image, video, audio, pdf
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.25
Output
$1.50
Cache read
$0.025
Cache write
$0.0833

Available from 1 host

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
edenai$0.25$1.50$0.025$0.08331M66K

More from Google

most-hosted first