Mistral model for multilingual chat, reasoning, and tool-assisted workflows
Specification
most-agreed values
Context
33K
Max output
33K
Released
2023-04-30
Knowledge cutoff
—
Retires
2025-11-13
Open weights
no
Input
text
Output
text
Price
US dollars per million tokens · most-agreed
Input
$0.20
Output
$0.20
Cache read
$0.20
Cache write
$0.20
Rate limits
Requests per minute
300
Available from 8 hosts
| Host | In $/M | Out $/M | Cache rd | Cache wr | Context | Output | Retires |
|---|---|---|---|---|---|---|---|
| nvidia | free | free | — | — | 66K | 66K | — |
| perplexity | $0.07 | $0.28 | — | — | 4K | 4K | — |
| openrouter | $0.13 | $0.13 | — | — | 33K | 8K | — |
| bedrock | $0.20 $0.15 | $0.26 $0.20 | — | — | 32K | 8K | — |
| fireworks_ai | $0.20 | $0.20 | — | — | 33K | 33K | — |
| pioneer | $0.20 | $0.20 | $0.20 | $0.20 | 33K | 33K | — |
| together_ai | $0.20 | $0.20 | — | — | 33K | — | 2025-11-13 |
| cloudflare | $1.923 | $1.923 | — | — | 8K | 8K | — |
Other listings of this model
same model, different key
More from Mistral
most-hosted first