modelbenchmark.io

GLM 5 Turbo

z-ai-glm-5-turbo

Compare

Faster GLM-5 lane for coding agents that need lower latency

Specification

most-agreed values

Context
200K
Max output
33K
Released
2026-03-15
Knowledge cutoff
Retires
Open weights
yes
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$1.20
Output
$4.00
Cache read
$0.24

Available from 1 host

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
venice$1.20$4.00$0.24200K33K

Other listings of this model

this page is one of them

The corpus files GLM 5 Turbo under several keys. This page is the z-ai-glm-5-turbo listing. The full record — every host, every price and every benchmark score — is on the main GLM 5 Turbo page.