modelbenchmark.io

Llama 3.2 11b Vision Instruct

Meta · llama-3-2-11b-vision-instruct

Compare

Open Llama multimodal model for image understanding and text reasoning

Specification

most-agreed values

Context
128K
Max output
4K
Released
2024-09-18
Knowledge cutoff
2023-12
Retires
2026-06-13
Open weights
yes
Input
text, image
Output
text

Price

US dollars per million tokens · most-agreed

Input
free
Output
free

Rate limits

Requests per minute
300

Available from 9 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
nvidiafreefree128K4K
cloudflare$0.0485$0.676128K128K
cloudflare-workers-ai · cf$0.0485$0.676128K128K
deepinfra$0.049$0.049131K131K
inference$0.055$0.05516K4K
edenai$0.345$0.345131K4K
watsonx$0.35$0.35128K128K
azure_ai$0.37$0.37128K2K2026-06-13
oci · oci$2.00$2.00128K4K

Other listings of this model

same model, different key

More from Meta

most-hosted first