modelbenchmark.io

Llama 3.1 8B

Meta · llama-3-1-8b-instant

Compare

Compact Llama instruction model for fast chat and local deployment

Specification

most-agreed values

Context
131K
Max output
131K
Released
2024-07-23
Knowledge cutoff
2023-12
Retires
2026-08-16
Open weights
yes
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.05
Output
$0.08

Available from 2 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
groq$0.05$0.08131K131K2026-08-16
helicone$0.05$0.08131K33K

Other listings of this model

this page is one of them

The corpus files Llama 3.1 8B under several keys. This page is the llama-3-1-8b-instant listing. The full record — every host, every price and every benchmark score — is on the main Llama 3.1 8B page.

More from Meta

most-hosted first