modelbenchmark.io

nemotron-mini-4b-instruct

NVIDIA · nemotron-mini-4b-instruct

Compare

Compact Nemotron model for efficient reasoning and deployable AI agents

Specification

most-agreed values

Context
128K
Max output
8K
Released
2024-08-21
Knowledge cutoff
Retires
Open weights
yes
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
free
Output
free

Available from 1 host

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
nvidiafreefree128K8K

More from NVIDIA

most-hosted first