modelbenchmark.io

Gemini Flash-Lite Latest

Google · gemini-flash-lite

Compare

Fast Gemini model balancing multimodal reasoning, tool use, and cost

Specification

google listing

Context
1M
Max output
66K
Released
2026-07-21
Knowledge cutoff
2026-03
Retires
Open weights
no
Input
text, image, video, audio, pdf
Output
text

Price

US dollars per million tokens · google list price

Input
$0.30
Output
$2.50
Cache read
$0.03
Cache write
$0.0833

Rate limits

as listed

Requests per minute
15
Tokens per minute
250,000

Available from 6 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
googlevendor$0.30$2.50$0.031M66K
gemini$0.10$0.40$0.011M66K
google-vertex$0.25$1.50$0.0251M66K
merge-gateway$0.25$1.50$0.0251M66K
orcarouter$0.25$1.50$0.0251M66K
nano-gpt$0.30$2.50$0.03$0.08331M66K

More from Google

most-hosted first