modelbenchmark.io

Nemotron 3.5 Lightning 30B A3B

NVIDIA · nvidia-nemotron-3-5-lightning-30b-a3b-bf16

Compare

Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads

Specification

most-agreed values

Context
8K
Max output
4K
Released
2026-08-11
Knowledge cutoff
Retires
Open weights
yes
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.50
Output
$0.50
Cache read
$0.50
Cache write
$0.50

Available from 2 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
runinfra$0.05$0.15$0.01262K33K
pioneer$0.50$0.50$0.50$0.508K4K

Other listings of this model

this page is one of them

The corpus files Nemotron 3.5 Lightning 30B A3B under several keys. This page is the nvidia-nemotron-3-5-lightning-30b-a3b-bf16 listing. The full record — every host, every price and every benchmark score — is on the main Nemotron 3.5 Lightning 30B A3B page.

More from NVIDIA

most-hosted first