Compact Llama instruction model for fast chat and local deployment
Specification
most-agreed values
Context
128K
Max output
128K
Released
2024-07-23
Knowledge cutoff
2024-07
Retires
—
Open weights
no
Input
text
Output
text
Price
US dollars per million tokens · most-agreed
Input
$0.02
Output
$0.03
Available from 1 host
| Host | In $/M | Out $/M | Cache rd | Cache wr | Context | Output | Retires |
|---|---|---|---|---|---|---|---|
| helicone | $0.02 | $0.03 | — | — | 128K | 128K | — |
Other listings of this model
this page is one of them
The corpus files Meta Llama 3.1 8B Instruct Turbo under several keys. This page is the llama-3-1-8b-instruct-turbo listing. The full record — every host, every price and every benchmark score — is on the main Meta Llama 3.1 8B Instruct Turbo page.
More from Meta
most-hosted first