modelbenchmark.io

Nemotron Super

NVIDIA · nemotron-120b-a12b

Compare

Nemotron middle tier for collaborative agents and high-volume reasoning workloads

Specification

most-agreed values

Context
203K
Max output
203K
Released
2026-03-11
Knowledge cutoff
2026-02
Retires
Open weights
yes
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.30
Output
$0.75
Cache read
$0.06

Available from 1 host

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
baseten$0.30$0.75$0.06203K203K

More from NVIDIA

most-hosted first