modelbenchmark.io

Gemini 2.5 Flash

Google · gemini-2-5-flash

Compare

Fast Gemini workhorse for multimodal apps where latency and price matter

Specification

google listing

Context
1M
Max output
66K
Released
2025-06-17
Knowledge cutoff
2025-01
Retires
2026-10-02
Open weights
no
Input
text, image, audio, video, pdf
Output
text

Price

US dollars per million tokens · google list price

Input
$0.30
Output
$2.50
Cache read
$0.03
Cache write
$0.30

Batch and priority prices

US dollars per million tokens · per listing

PriceListing$/M
Batch inputvertex_ai-language-models$0.15
Batch outputvertex_ai-language-models$1.25
Priority inputvertex_ai-language-models$0.54
Priority outputvertex_ai-language-models$4.50

Batch input is 50% below the interactive input price at the same listing.

Rate limits

as listed

Requests per minute
100,000
Tokens per minute
8,000,000

Quality

2 benchmarks · 2 sources

Composite
40th
percentile of 334 scored models
Rank
200 / 334
±0.08 sd
Evidence
2 × 2
one source, or fewer than 3 benchmarks
Effort range
one configuration only
coding12th1/3 bench
math80th1/7 bench

Every score

one row per source and configuration — nothing averaged away

BenchmarkScoreConfigurationSourceRun
OTIS Mock AIME 2024-202571.9 ±5.7default · 2 runsEpoch AI2025-05-20
SWE-Bench verified28.7default · mini-SWE-agentSWE-bench

The leading figure for a benchmark is the median across its configurations, so one heroic high-effort run cannot set the number.

Available from 40 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
googlevendor$0.30$2.50$0.031M66K
kenarifreefree1M66K
qihang-ai$0.09$0.711M66K
oci · oci$0.15$0.601M66K
poe$0.21$1.80$0.0211.1M66K
jiekou$0.27$2.251M66K
cortecs$0.299$2.491$0.029$0.0971M66K
302ai$0.30$2.501M66K
abacus$0.30$2.50$0.031M66K
aihubmix$0.30$2.50$0.031M66K
auriko$0.30$2.50$0.031M66K
crossmodel$0.30$2.50$0.03$0.301M66K
databricks$0.30$2.50$0.03$0.301M
66K
66K
2026-10-02
deepinfra$0.30$2.501M1M
fastrouter$0.30$2.50$0.03751M66K
frogbot$0.30$2.50$0.0751M66K
gemini$0.30$2.50$0.031M66K
google-vertex$0.30$2.50$0.031M66K
helicone$0.30$2.50$0.075$0.301M66K
impossibl$0.30$2.50$0.031M66K
kilo$0.30$2.50$0.03$0.08331M66K
llmgateway$0.30$2.50$0.031M66K
llmgateway-providers · google-ai-studio$0.30$2.50$0.031M66K
llmgateway-providers · google-vertex$0.30$2.50$0.031M66K
merge-gateway$0.30$2.50$0.031M66K
modelis$0.30$2.501M66K
nano-gpt$0.30$2.50$0.031M66K
nearai$0.30$2.50$0.031M66K
ofox$0.30$2.50$0.03$1.001M66K
openrouter$0.30$2.50$0.03$0.08331M66K
orcarouter$0.30$2.50$0.031M66K
perplexity-agent$0.30$2.50$0.031M66K
sap-ai-core$0.30$2.50$0.031M66K
vercel$0.30$2.50$0.031M66K
vercel_ai_gateway · vercel-ai-gateway$0.30$2.501M66K
vertex_ai-language-models$0.30$2.50$0.031M66K2026-10-20
zenmux$0.30$2.50$0.07$1.001M64K
replicate · replicate$2.50$2.50
anyapi1M66K
qiniu-ai1M64K

Price history

append-only observations · a listing writes a row only when its price moves

DateHostField$/M
2026-09-02openrouterCache read$0.03
2026-09-06openrouterCache read$0.03

More from Google

most-hosted first