modelbenchmark.io

NVIDIA Nemotron 3 Nano 30B

NVIDIA · nvidia-nemotron-3-nano-30b-a3b

Compare

Small Nemotron 3 MoE for efficient coding, math, and long-context agents

Specification

most-agreed values

Context
128K
Max output
16K
Released
2026-01-27
Knowledge cutoff
Retires
Open weights
yes
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.075
Output
$0.30
Cache read
$0.03

Available from 4 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
crusoe$0.05$0.20$0.03262K262K
cortecs$0.06$0.24256K256K
nebius$0.06$0.24262K262K
venice$0.075$0.30128K16K

Price history

append-only observations · a listing writes a row only when its price moves

DateHostField$/M
2026-09-02nebiusInput$0.06
2026-09-06nebiusInput$0.06
2026-09-02nebiusOutput$0.24
2026-09-06nebiusOutput$0.24

Other listings of this model

this page is one of them

The corpus files NVIDIA Nemotron 3 Nano 30B under several keys. This page is the nvidia-nemotron-3-nano-30b-a3b listing. The full record — every host, every price and every benchmark score — is on the main NVIDIA Nemotron 3 Nano 30B page.

More from NVIDIA

most-hosted first