modelbenchmark.io

Nemotron 3 Super 120B A12B

NVIDIA · nvidia-nemotron-3-super-120b-a12b-fp8

Compare

Nemotron middle tier for collaborative agents and high-volume reasoning workloads

Specification

most-agreed values

Context
256K
Max output
32K
Released
2026-03-11
Knowledge cutoff
Retires
Open weights
yes
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.09
Output
$0.45
Cache read
$0.09
Cache write
$0.09

Available from 1 host

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
pioneer$0.09$0.45$0.09$0.09256K32K

More from NVIDIA

most-hosted first