Nemotron model for efficient reasoning, coding, and specialized AI agents
Specification
most-agreed values
Context
131K
Max output
131K
Released
2025-08-08
Knowledge cutoff
—
Retires
—
Open weights
yes
Input
text
Output
text
Price
US dollars per million tokens · most-agreed
Input
$0.15
Output
$0.40
Cache read
$0.075
Available from 4 hosts
More from NVIDIA
most-hosted first