modelbenchmark.io

Mercury 2

mercury-2

Compare

Compact GPT model for low-latency assistance and high-volume workloads

Specification

most-agreed values

Context
128K
Max output
50K
Released
2024-01-01
Knowledge cutoff
2025-01-01
Retires
Open weights
no
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.25
Output
$0.75
Cache read
$0.025

Available from 7 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
inception$0.25$0.75$0.025128K50K
inception · inception$0.25$0.75$0.025128K50K
kilo · inception$0.25$0.75$0.025128K50K
nano-gpt$0.25$0.75$0.025128K50K
openrouter · inception$0.25$0.75$0.025128K50K
vercel · inception$0.25$0.75$0.025128K128K
venice$0.3125$0.9375$0.0312128K50K