modelbenchmark.io

Phi 4 Multimodal

Microsoft · phi-4-multimodal-instruct

Compare

Compact GPT model for low-latency assistance and high-volume workloads

Specification

most-agreed values

Context
128K
Max output
16K
Released
2025-07-26
Knowledge cutoff
Retires
Open weights
yes
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.07
Output
$0.11
Cache read
$0.035

Available from 3 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
nvidiafreefree128K16K
nano-gpt$0.07$0.11$0.035128K16K
azure_ai$0.08$0.32131K4K

Other listings of this model

same model, different key

More from Microsoft

most-hosted first