modelbenchmark.io

OpenAI: o1-mini

OpenAI · o1-mini

Compare

O-series reasoning model for hard analysis, math, coding, and planning

Specification

most-agreed values

Context
128K
Max output
66K
Released
2025-01-01
Knowledge cutoff
2025-01
Retires
Open weights
no
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$1.10
Output
$4.40
Cache read
$0.55

Batch and priority prices

US dollars per million tokens · per listing

PriceListing$/M
Batch inputazure$0.605
Batch outputazure$2.42

Batch input is 50% below the interactive input price at the same listing.

Quality

5 benchmarks · 2 sources

Composite
49th
percentile of 334 scored models
Rank
172 / 334
±0.05 sd
Evidence
5 × 2
3 or more benchmarks, from 2 or more sources
Effort range
one configuration only
coding26th1/3 bench
reasoning53rd1/4 bench
math55th3/7 bench

Every score

one row per source and configuration — nothing averaged away

BenchmarkScoreConfigurationSourceRun
MATH level 586.7 ±0.7default · 2 runsEpoch AI2025-02-13
GPQA diamond60.9 ±2.7default · 2 runsEpoch AI2025-02-13
OTIS Mock AIME 2024-202545.8 ±6.5default · 2 runsEpoch AI2025-03-06
Aider polyglot32.9default · aider wholeAider2024-12-22
FrontierMath-v10.8 ±0.0default · 4 runsEpoch AI2025-03-06

The leading figure for a benchmark is the median across its configurations, so one heroic high-effort run cannot set the number.

Available from 3 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
helicone$1.10$4.40$0.55128K66K
replicate · replicate$1.10$4.40
azure
$1.21
$1.10
$4.84
$4.40
$0.605
$0.55
128K66K

More from OpenAI

most-hosted first