modelbenchmark.io

Gemini 3.5 Flash Lite

Google · gemini-3-5-flash-lite

Compare

Fast Gemini model balancing multimodal reasoning, tool use, and cost

Specification

google listing

Context
1M
Max output
66K
Released
2026-07-21
Knowledge cutoff
2026-03
Retires
2027-07-21
Open weights
no
Input
text, image, video, audio, pdf
Output
text

Price

US dollars per million tokens · google list price

Input
$0.30
Output
$2.50
Cache read
$0.03
Cache write
$0.0833

Batch and priority prices

US dollars per million tokens · per listing

PriceListing$/M
Batch inputvertex_ai-language-models$0.15
Batch outputvertex_ai-language-models$1.25
Priority inputvertex_ai-language-models$0.54
Priority outputvertex_ai-language-models$4.50

Batch input is 50% below the interactive input price at the same listing.

Rate limits

as listed

Requests per minute
15
Tokens per minute
250,000

Quality

7 benchmarks · 2 sources

Composite
39th
percentile of 334 scored models
Rank
204 / 334
±0.09 sd
Evidence
7 × 2
3 or more benchmarks, from 2 or more sources
Effort range
8.0
points between effort settings
reasoning69th3/4 bench
math32nd3/7 bench
general2nd1/1 bench

By reasoning effort

every setting placed on the same scale as the leaderboard

SettingCompositeBenchmarks behind it
low65th4
minimal59th4
high43rd6

Every score

one row per source and configuration — nothing averaged away

BenchmarkScoreConfigurationSourceRun
GPQA diamond83.3 ±2.7highEpoch AI2026-08-06
75.8 ±3.1lowEpoch AI2026-08-06
74.2 ±3.1minimalEpoch AI2026-08-06
Settings differ by 9.1 points on this benchmark.
Chess Puzzles22.0 ±4.2highEpoch AI2026-08-06
21.0 ±4.1minimalEpoch AI2026-08-06
18.0 ±3.9lowEpoch AI2026-08-06
Settings differ by 4.0 points on this benchmark.
OTIS Mock AIME 2024-202571.1 ±6.8highEpoch AI2026-08-06
60.0 ±7.4lowEpoch AI2026-08-06
51.1 ±7.5minimalEpoch AI2026-08-06
Settings differ by 20.0 points on this benchmark.
Mystery Game Puzzles19.0 ±3.9lowEpoch AI2026-08-05
12.0 ±3.3minimalEpoch AI2026-08-05
Settings differ by 7.0 points on this benchmark.
FrontierMath-Tiers-1-3-v226.0 ±2.6highEpoch AI2026-08-02
FrontierMath-Tier-4-v20.0 ±0.0highEpoch AI2026-08-02
LiveBench63.8high · 23 runsLiveBench2026-06-25

The leading figure for a benchmark is the median across its configurations, so one heroic high-effort run cannot set the number.

Available from 29 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
googlevendor$0.30$2.50$0.031M66K
kilo$0.15$1.25$0.015$0.04171M66K
302ai$0.30$2.501M66K
abacus$0.30$2.50$0.031M66K
crossmodel$0.30$2.50$0.03$0.301M66K
edenai$0.30$2.50$0.03$0.08331M66K
gemini$0.30$2.50$0.031M66K
google-vertex$0.30$2.50$0.031M66K
impossibl$0.30$2.50$0.031M66K
llmgateway$0.30$2.50$0.03$0.08331M66K
llmgateway-providers · google-ai-studio$0.30$2.50$0.03$0.08331M66K
llmgateway-providers · google-vertex$0.30$2.50$0.03$0.08331M66K
merge-gateway$0.30$2.50$0.031M66K
nano-gpt$0.30$2.50$0.03$0.08331M66K
neon$0.30$2.50$0.031M66K
ofox$0.30$2.50$0.03$0.0831M66K
opencode$0.30$2.50$0.031M66K
openrouter$0.30$2.50$0.03$0.08331M66K
opper$0.30$2.50$0.031M66K
orcarouter$0.30$2.50$0.031M66K
perplexity$0.30$2.50$0.03
pioneer$0.30$2.50$0.03$0.301M65K
requesty$0.30$2.50$0.031M66K
vercel$0.30$2.50$0.031M65K
vertex_ai-language-models$0.30$2.50$0.031M66K2027-07-21
cortecs$0.33$2.749$0.0331M66K
databricks$0.375$3.125$0.0375$0.3751M66K
venice$0.375$3.125$0.03751M66K
sap-ai-core1M66K

Price history

append-only observations · a listing writes a row only when its price moves

DateHostField$/M
2026-09-02openrouterCache read$0.03
2026-09-06openrouterCache read$0.03
2026-09-02openrouterInput$0.30
2026-09-06openrouterInput$0.30
2026-09-02openrouterOutput$2.50
2026-09-06openrouterOutput$2.50

More from Google

most-hosted first