modelbenchmark.io

GLM-5.3 Flash Flex

Zhipu · glm-5-3-flash-flex

Compare

Native multimodal GLM model for efficient coding and long-horizon agent tasks

Specification

most-agreed values

Context
1M
Max output
1M
Released
2026-08-26
Knowledge cutoff
—
Retires
—
Open weights
yes
Input
text, image
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.0975
Output
$0.325
Cache read
$0.0195

Available from 1 host

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
neuralwatt$0.0975$0.325$0.0195—1M1M—

More from Zhipu

most-hosted first