modelbenchmark.io

Nvidia Nemotron 70b

NVIDIA · llama-3-1-nemotron-70b-instruct-hf

Compare

Nemotron model for efficient reasoning, coding, and specialized AI agents

Specification

most-agreed values

Context
16K
Max output
8K
Released
2025-04-15
Knowledge cutoff
Retires
Open weights
yes
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.357
Output
$0.408
Cache read
$0.1785

Available from 2 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
nano-gpt$0.357$0.408$0.178516K8K
together_ai$0.88$0.88

More from NVIDIA

most-hosted first