modelbenchmark.io

Gemma 4 E4B Instruct

Google · gemma-4-e4b-it

Compare

Open Gemma instruction model for efficient chat and self-hosted deployments

Specification

most-agreed values

Context
131K
Max output
16K
Released
2026-04-02
Knowledge cutoff
Retires
Open weights
yes
Input
text, image, video, audio
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.02
Output
$0.10
Cache read
$0.02
Cache write
$0.20

Available from 3 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
deepinfra$0.02$0.10131K8K
nano-gpt$0.04$0.20$0.02131K16K
pioneer$0.20$0.20$0.20$0.2033K33K

More from Google

most-hosted first