Price against quality
The frontier
Every scored model, plotted against what it costs to run. The models on the step are the ones nothing else beats on both price and score. Read it left to right: the cheapest model that clears the quality you need is the one you want.
Overall percentile against output price
73 models plotted · 8 on the frontier · firm evidence only
- On the frontier
- Beaten on both
- Open weights
Output price, $ per million tokens · log scale →↑ Overall percentile
Cheapest on the frontier
Best score at any price
On the frontier
| Model | Lab | Composite | Out $/M | In $/M | Evidence |
|---|---|---|---|---|---|
| GPT-6 Astra | OpenAI | 100th | $50.00 | $10.00 | 11 bench · 2 src |
| Claude Opus 5 | Anthropic | 96th | $25.00 | $5.00 | 9 bench · 2 src |
| Gemini 3.8 Flash | 95th | $3.75 | $0.75 | 8 bench · 2 src | |
| Gemini 3 Flash Preview | 84th | $3.00 | $0.50 | 10 bench · 2 src | |
| DeepSeek V4 Flash 0731 | DeepSeek | 81st | $0.50 | $0.20 | 8 bench · 2 src |
| GPT-5.6 Luna | OpenAI | 80th | $0.37 | $0.06 | 8 bench · 2 src |
| DeepSeek-V3.2-Exp | DeepSeek | 68th | $0.326 | $0.2174 | 7 bench · 3 src |
| GLM-5.3-Flash | Zhipu | 61st | $0.025 | $0.075 | 7 bench · 2 src |
Price is the catalog list price for output tokens, from the model's own vendor where it has a listing. A model with no price in the catalog cannot be plotted. See the full ranking.