modelbenchmark.io

GPT 5.1 Thinking (Fast)

OpenAI · gpt-5-1-thinking-fast

Compare

Compact GPT model for low-latency assistance and high-volume workloads

Specification

most-agreed values

Context
400K
Max output
128K
Released
2025-11-12
Knowledge cutoff
Retires
Open weights
no
Input
text, image, pdf
Output
text

Price

US dollars per million tokens · most-agreed

Input
$2.50
Output
$20.00
Cache read
$0.25

Available from 1 host

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
vercel$2.50$20.00$0.25400K128K

More from OpenAI

most-hosted first