GLM 5 Turbo
z-ai-glm-5-turbo
Faster GLM-5 lane for coding agents that need lower latency
Specification
most-agreed values
Context
200K
Max output
33K
Released
2026-03-15
Knowledge cutoff
—
Retires
—
Open weights
yes
Input
text
Output
text
Price
US dollars per million tokens · most-agreed
Input
$1.20
Output
$4.00
Cache read
$0.24
Available from 1 host
| Host | In $/M | Out $/M | Cache rd | Cache wr | Context | Output | Retires |
|---|---|---|---|---|---|---|---|
| venice | $1.20 | $4.00 | $0.24 | — | 200K | 33K | — |
Other listings of this model
this page is one of them
The corpus files GLM 5 Turbo under several keys. This page is the z-ai-glm-5-turbo listing. The full record — every host, every price and every benchmark score — is on the main GLM 5 Turbo page.