Native multimodal GLM model for efficient coding and long-horizon agent tasks
Specification
most-agreed values
Context
1M
Max output
131K
Released
2026-08-26
Knowledge cutoff
—
Retires
—
Open weights
yes
Input
text, image, video, pdf
Output
text
Price
US dollars per million tokens · most-agreed
Input
$0.1187
Output
$0.4156
Cache read
$0.0341
Available from 1 host
| Host | In $/M | Out $/M | Cache rd | Cache wr | Context | Output | Retires |
|---|---|---|---|---|---|---|---|
| volcengine | $0.1187 | $0.4156 | $0.0341 | — | 1M | 131K | — |
Other listings of this model
this page is one of them
The corpus files GLM-5.3-Flash under several keys. This page is the glm-5-3-flash-260828 listing. The full record — every host, every price and every benchmark score — is on the main GLM-5.3-Flash page.
More from Zhipu
most-hosted first