modelbenchmark.io

Gemini 3.6 Flash

Google · gemini-3-6-flash

Compare

Fast Gemini model balancing multimodal reasoning, tool use, and cost

Specification

google listing

Context
1M
Max output
66K
Released
2026-07-21
Knowledge cutoff
2026-03
Retires
Open weights
no
Input
text, image, video, audio, pdf
Output
text

Price

US dollars per million tokens · google list price

Input
$0.75
Output
$3.75
Cache read
$0.075
Cache write
$0.0417

Batch and priority prices

US dollars per million tokens · per listing

PriceListing$/M
Batch inputvertex_ai$0.375
Batch outputvertex_ai$1.875
Priority inputvertex_ai$1.35
Priority outputvertex_ai$6.75

Batch input is 50% below the interactive input price at the same listing.

Rate limits

as listed

Requests per minute
2,000
Tokens per minute
800,000

Quality

8 benchmarks · 2 sources

Composite
84th
percentile of 334 scored models
Rank
54 / 334
±0.08 sd
Evidence
8 × 2
3 or more benchmarks, from 2 or more sources
Effort range
8.1
points between effort settings
reasoning92nd3/4 bench
math71st3/7 bench
knowledge89th1/1 bench
general39th1/1 bench

By reasoning effort

every setting placed on the same scale as the leaderboard

SettingCompositeBenchmarks behind it
low89th4
minimal88th4
high87th8

Every score

one row per source and configuration — nothing averaged away

BenchmarkScoreConfigurationSourceRun
Chess Puzzles43.0 ±5.0lowEpoch AI2026-08-07
40.0 ±4.9highEpoch AI2026-08-02
35.0 ±4.8minimalEpoch AI2026-08-07
Settings differ by 8.0 points on this benchmark.
SimpleQA Verified66.2 ±1.5highEpoch AI2026-08-27
GPQA diamond94.1 ±1.4highEpoch AI2026-08-02
86.4 ±2.4lowEpoch AI2026-08-06
85.9 ±2.5minimalEpoch AI2026-08-07
Settings differ by 8.2 points on this benchmark.
OTIS Mock AIME 2024-202594.2 ±3.1highEpoch AI2026-08-02
82.2 ±5.8lowEpoch AI2026-08-07
80.0 ±6.0minimalEpoch AI2026-08-07
Settings differ by 14.2 points on this benchmark.
FrontierMath-Tiers-1-3-v259.0 ±2.9highEpoch AI2026-08-02
Mystery Game Puzzles30.0 ±4.6highEpoch AI2026-08-05
25.0 ±4.4minimalEpoch AI2026-08-05
23.0 ±4.2lowEpoch AI2026-08-05
Settings differ by 7.0 points on this benchmark.
LiveBench74.5high · 23 runsLiveBench2026-06-25
FrontierMath-Tier-4-v221.9 ±6.5highEpoch AI2026-08-02

The leading figure for a benchmark is the median across its configurations, so one heroic high-effort run cannot set the number.

Available from 30 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
googlevendor$0.75$3.75$0.0751M66K
kenarifreefree1M66K
kilo$0.375$1.875$0.0375$0.02081M66K
cortecs$0.75$3.75$0.075$0.0381M66K
crossmodel$0.75$3.75$0.075$0.751M66K
edenai$0.75$3.75$0.075$0.04171M66K
gemini$0.75$3.75$0.0751M66K
github-copilot$0.75$3.75$0.0751M64K
google-vertex$0.75$3.75$0.0751M66K
llmgateway$0.75$3.75$0.075$0.08331M66K
llmgateway-providers · google-ai-studio$0.75$3.75$0.075$0.08331M66K
llmgateway-providers · google-vertex$0.75$3.75$0.075$0.08331M66K
nano-gpt$0.75$3.75$0.075$0.04171M66K
ofox$0.75$3.75$0.075$0.04151M66K
openrouter$0.75$3.75$0.075$0.04171M66K
vercel$0.75$3.75$0.0751M64K
vertex_ai$0.75$3.75$0.0751M66K
vertex_ai-language-models$0.75$3.75$0.0751M66K
venice$0.9375$4.6875$0.09381M66K
302ai$1.50$7.501M66K
abacus$1.50$7.50$0.151M66K
impossibl$1.50$7.50$0.151M66K
merge-gateway$1.50$7.50$0.151M66K
neon$1.50$7.50$0.151M66K
opencode$1.50$7.50$0.151M66K
orcarouter$1.50$7.50$0.151M66K
perplexity$1.50$7.50$0.15
pioneer$1.50$7.50$0.15$1.501M64K
requesty$1.50$7.00$0.151M66K
databricks$1.875$9.375$0.1875$1.8751M66K

Price history

append-only observations · a listing writes a row only when its price moves

DateHostField$/M
2026-09-02crossmodelCache read$0.15
2026-09-05crossmodelCache read$0.075
2026-09-02openrouterCache read$0.075
2026-09-06openrouterCache read$0.075
2026-09-02crossmodelCache write$1.50
2026-09-05crossmodelCache write$0.75
2026-09-02crossmodelInput$1.50
2026-09-05crossmodelInput$0.75
2026-09-02openrouterInput$0.75
2026-09-06openrouterInput$0.75
2026-09-02crossmodelOutput$7.50
2026-09-05crossmodelOutput$3.75
2026-09-02openrouterOutput$3.75
2026-09-06openrouterOutput$3.75

More from Google

most-hosted first