modelbenchmark.io

Llama 4 Maverick 17B 128E Instruct

Meta · llama-4-maverick-17b-128e-instruct-maas

Compare

Open multimodal Llama model for strong reasoning and fast responses

Specification

most-agreed values

Context
524K
Max output
8K
Released
2025-04-29
Knowledge cutoff
2024-08
Retires
Open weights
yes
Input
text, image
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.35
Output
$1.15

Available from 2 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
google-vertex$0.35$1.15524K8K
vertex_ai-llama_models$0.35$1.151M1M

More from Meta

most-hosted first