modelbenchmark.io

Llama 3.1 Nemotron 70B Instruct

NVIDIA · llama-3-1-nemotron-70b-instruct

Compare

Nemotron model for efficient reasoning, coding, and specialized AI agents

Specification

most-agreed values

Context
128K
Max output
8K
Released
2025-04-15
Knowledge cutoff
Retires
Open weights
yes
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
free
Output
free

Available from 3 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
nvidiafreefree128K8K
deepinfra$0.60$0.60131K131K
edenai$0.60$0.60131K8K

More from NVIDIA

most-hosted first