modelbenchmark.io

Claude 4.1 Opus

Anthropic · claude-opus-4-1

Compare

Flagship Claude model for deep reasoning, coding, and long-horizon agents

Specification

anthropic listing

Context
200K
Max output
32K
Released
2025-08-05
Knowledge cutoff
2025-03-31
Retires
2026-08-05
Open weights
no
Input
text, image, pdf
Output
text

Price

US dollars per million tokens · anthropic list price

Input
$15.00
Output
$75.00
Cache read
$1.50
Cache write
$18.75

Batch and priority prices

US dollars per million tokens · per listing

PriceListing$/M
Batch inputvertex_ai-anthropic_models$7.50
Batch outputvertex_ai-anthropic_models$37.50
Cache write, 1 hourazure_ai$30.00

Batch input is 50% below the interactive input price at the same listing.

Quality

10 benchmarks · 1 source

Composite
43rd
percentile of 334 scored models
Rank
190 / 334
±0.06 sd
Evidence
10 × 1
one source, or fewer than 3 benchmarks
Effort range
4.0
points between effort settings
coding77th1/3 bench
reasoning49th4/4 bench
math30th5/7 bench

By reasoning effort

every setting placed on the same scale as the leaderboard

SettingCompositeBenchmarks behind it
24k58th1
16k57th3
27k57th4
32k16th2

Every score

one row per source and configuration — nothing averaged away

BenchmarkScoreConfigurationSourceRun
SWE-Bench verified73.3 ±2.0defaultEpoch AI2026-02-11
GPQA diamond77.3 ±3.016kEpoch AI2025-08-05
76.8 ±3.027kEpoch AI2025-08-05
73.2 ±3.2defaultEpoch AI2025-08-05
Settings differ by 4.1 points on this benchmark.
OTIS Mock AIME 2024-202568.9 ±7.027kEpoch AI2025-08-05
64.4 ±7.216kEpoch AI2025-08-05
40.0 ±7.4defaultEpoch AI2025-08-05
Settings differ by 28.9 points on this benchmark.
Mystery Game Puzzles21.0 ±4.124kEpoch AI2026-07-25
FrontierMath-Tier-4-2025-07-014.2 ±2.927kEpoch AI2025-08-05
Chess Puzzles7.0 ±2.6defaultEpoch AI2026-07-20
FrontierMath-v13.6 ±1.527k · 2 runsEpoch AI2025-08-05
2.9 ±1.4default · 2 runsEpoch AI2025-08-05
0.0 ±0.016kEpoch AI2025-08-05
Settings differ by 3.6 points on this benchmark.
EBR-bench7.9defaultEpoch AI2026-06-25
FrontierMath-Tier-4-v22.4 ±2.432kEpoch AI2026-06-11
FrontierMath-Tiers-1-3-v212.6 ±2.032kEpoch AI2026-06-11

The leading figure for a benchmark is the median across its configurations, so one heroic high-effort run cannot set the number.

Available from 32 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
anthropicvendor$15.00$75.00$1.50$18.75200K32K2026-08-05
poe$13.00$64.00$1.30$16.00197K32K
jiekou$13.50$67.50200K32K
302ai$15.00$75.00200K32K
abacus$15.00$75.00200K32K
amazon-bedrock$15.00$75.00$1.50$18.75200K32K
amazon-bedrock · us$15.00$75.00$1.50$18.75200K32K
azure$15.00$75.00$1.50$18.75200K32K
azure_ai$15.00$75.00$1.50$18.75200K32K2026-08-05
azure-cognitive-services$15.00$75.00$1.50$18.75200K32K
bedrock_converse$15.00$75.00$1.50$18.75200K32K2027-01-08
bedrock_converse · eu$15.00$75.00$1.50$18.75200K32K2027-01-08
bedrock_converse · us$15.00$75.00$1.50$18.75200K32K2027-01-08
databricks$15.00$75.00$1.50$18.75200K32K
fastrouter$15.00$75.00$1.50$18.75200K32K
google-vertex$15.00$75.00$1.50$18.75200K32K
google-vertex-anthropic$15.00$75.00$1.50$18.75200K32K
helicone$15.00$75.00$1.50$18.75200K32K
kilo$15.00$75.00$1.50$18.75200K32K
llmgateway$15.00$75.00$1.50$18.75200K32K
llmgateway-providers · aws-bedrock$15.00$75.00$1.50$18.75200K32K
merge-gateway$15.00$75.00$1.50$18.75200K32K
nano-gpt$15.00$75.00$1.50200K32K
nano-gpt · thinking$15.00$75.00$1.50200K32K
neon$15.00$75.00$1.50$18.75200K32K
opencode$15.00$75.00$1.50$18.75200K32K
openrouter$15.00$75.00$1.50$18.75200K32K
pioneer$15.00$75.00$1.50$18.75200K32K
requesty$15.00$75.00$1.50$18.75200K32K
vercel_ai_gateway · vercel-ai-gateway$15.00$75.00$1.50$18.75200K32K
vertex_ai-anthropic_models$15.00$75.00$1.50$18.75200K32K2026-08-05
zenmux$15.00$75.00$1.50$18.75200K64K

Other listings of this model

same model, different key

More from Anthropic

most-hosted first