modelbenchmark.io

Nemotron 3 Nano 30B A3B FP8

NVIDIA · nvidia-nemotron-3-nano-30b-a3b-fp8

Compare

Small Nemotron 3 MoE for efficient coding, math, and long-context agents

Specification

most-agreed values

Context
1M
Max output
262K
Released
2025-12-15
Knowledge cutoff
Retires
Open weights
yes
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.06
Output
$0.25

Available from 1 host

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
infomaniak$0.06$0.251M262K

More from NVIDIA

most-hosted first