modelbenchmark.io

GLM5.2-Fast

glm5-2-fast

Compare

The same model served for high TPS.

Specification

most-agreed values

Context
1M
Max output
131K
Released
2026-06-13
Knowledge cutoff
Retires
Open weights
yes
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$3.00
Output
$10.25
Cache read
$0.50
Cache write
free

Available from 1 host

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
wafer.ai$3.00$10.25$0.50free1M131K