modelbenchmark.io

Gemini 2.5 Flash Preview Thinking

Google · gemini-2-5-flash-preview-04-17

Compare

Compact GPT model for low-latency assistance and high-volume workloads

Specification

most-agreed values

Context
1M
Max output
66K
Released
2025-04-17
Knowledge cutoff
Retires
Open weights
no
Input
text, image, audio
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.15
Output
$3.50
Cache read
$0.015

Available from 2 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
nano-gpt$0.15$0.60$0.0151M66K
nano-gpt · thinking$0.15$3.50$0.0151M66K

More from Google

most-hosted first