modelbenchmark.io

Llama-4-Scout-17B-16E-Instruct-FP8

Meta · llama-4-scout-17b-16e-instruct-fp8

Compare

Open multimodal Llama model for long-context analysis and efficient agents

Specification

most-agreed values

Context
128K
Max output
4K
Released
2025-04-05
Knowledge cutoff
2024-08
Retires
Open weights
yes
Input
text, image
Output
text

Price

US dollars per million tokens · most-agreed

Input
free
Output
free

Available from 2 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
llamafreefree128K4K
meta_llama · meta-llama10M4K

More from Meta

most-hosted first