modelbenchmark.io

Grok 4.6

xAI · grok-4-6

Compare

xAI's frontier model for long-running agents, coding, knowledge work, and visual projects

Specification

xai listing

Context
500K
Max output
500K
Released
2026-08-12
Knowledge cutoff
2026-02-01
Retires
Open weights
no
Input
text, image, pdf
Output
text

Price

US dollars per million tokens · xai list price

Input
$2.00
Output
$6.00
Cache read
$0.50
Cache write
$2.00

Long-context price tiers

the rate above each context threshold

AbovePrice$/M
200KCache read$1.00
200KInput$4.00
200KOutput$12.00

Quality

9 benchmarks · 2 sources

Composite
87th
percentile of 334 scored models
Rank
45 / 334
±0.07 sd
Evidence
9 × 2
3 or more benchmarks, from 2 or more sources
Effort range
1.1
points between effort settings
reasoning92nd4/4 bench
math86th3/7 bench
knowledge69th1/1 bench
general79th1/1 bench

By reasoning effort

every setting placed on the same scale as the leaderboard

SettingCompositeBenchmarks behind it
high93rd4
xhigh87th8

Every score

one row per source and configuration — nothing averaged away

BenchmarkScoreConfigurationSourceRun
OTIS Mock AIME 2024-202599.2 ±0.5xhighEpoch AI2026-08-14
97.8 ±1.0highEpoch AI2026-08-12
Settings differ by 1.4 points on this benchmark.
GPQA diamond94.0 ±1.4highEpoch AI2026-08-12
93.2 ±1.5xhighEpoch AI2026-08-14
Settings differ by 0.8 points on this benchmark.
Chess Puzzles40.0 ±4.9highEpoch AI2026-08-12
31.0 ±4.6xhighEpoch AI2026-08-14
Settings differ by 9.0 points on this benchmark.
Mystery Game Puzzles34.0 ±4.8xhighEpoch AI2026-08-14
FrontierMath-Tiers-1-3-v266.0 ±2.8xhighEpoch AI2026-08-14
LiveBench79.0default · 23 runsLiveBench2026-06-25
SimpleQA Verified49.3 ±1.6highEpoch AI2026-08-27
48.9 ±1.6xhighEpoch AI2026-08-27
Settings differ by 0.4 points on this benchmark.
EBR-bench30.5 ±1.6xhighEpoch AI2026-08-18
FrontierMath-Tier-4-v231.7 ±7.4xhighEpoch AI2026-08-14

The leading figure for a benchmark is the median across its configurations, so one heroic high-effort run cannot set the number.

Available from 40 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
xaivendor$2.00$6.00$0.50500K500K
kenarifreefree500K500K
302ai$2.00$6.00500K500K
abacus$2.00$6.00$0.50500K33K
aihubmix$2.00$6.00$0.50500K500K
amazon-bedrock · global$2.00$6.00$0.50500K500K
azure_ai$2.00$6.00$0.50200K128K
bedrock_converse · global$2.00$6.00$0.50500K500K
cloudflare-ai-gateway$2.00$6.00$0.50500K500K
crossmodel · x-ai$2.00$6.00$0.50$2.00500K500K
edenai$2.00$6.00$0.50500K500K
github-copilot$2.00$6.00$0.50500K128K
google-vertex$2.00$6.00$0.50524K500K
kilo · x-ai$2.00$6.00$0.50500K450K
llmgateway$2.00$6.00$0.50500K500K
llmgateway-providers$2.00$6.00$0.50500K500K
llmgateway-providers · aws-bedrock$2.00$6.00$0.50500K500K
llmgateway-providers · vertex-openai$2.00$6.00$0.50500K500K
merge-gateway$2.00$6.00$0.50500K500K
nano-gpt · x-ai$2.00$6.00$0.50500K450K
ofox · x-ai$2.00$6.00$0.50500K66K
opencode$2.00$6.00$0.50500K500K
opencode-go$2.00$6.00$0.50500K500K
openrouter$2.00$6.00$0.50500K500K
openrouter · x-ai$2.00$6.00$0.50500K450K
opper$2.00$6.00$0.50500K500K
orcarouter · grok$2.00$6.00$0.50500K500K
perplexity$2.00$6.00$0.50
perplexity-agent$2.00$6.00$0.50500K500K
requesty$2.00$6.00$0.50$2.00500K500K
vercel · spacexai$2.00$6.00$0.50500K500K
vertex_ai$2.00$6.00$0.50524K524K
amazon-bedrock$2.20$6.60$0.55500K500K
amazon-bedrock · us$2.20$6.60$0.55500K500K
bedrock_converse · us$2.20$6.60$0.55500K500K
bedrock_mantle · bedrock-mantle
$2.20
$2.64
$6.60
$7.92
$0.55
$0.66
500K500K
venice$2.27$6.80$0.57500K200K
databricks$2.50$7.50$0.625$2.50500K
azure200K128K
neon500K524K

Price history

append-only observations · a listing writes a row only when its price moves

DateHostField$/M
2026-09-02bedrock_mantle · bedrock-mantleCache read$0.55
2026-09-09bedrock_mantle · bedrock-mantleCache read$0.55
2026-09-02bedrock_mantle · bedrock-mantleInput$2.20
2026-09-09bedrock_mantle · bedrock-mantleInput$2.20
2026-09-02bedrock_mantle · bedrock-mantleOutput$6.60
2026-09-09bedrock_mantle · bedrock-mantleOutput$6.60

More from xAI

most-hosted first