modelbenchmark.io

GLM 5.2 Short Fast

Zhipu · glm-5-2-short-fast

Compare

Efficient GLM model for fast reasoning, coding, and agent workflows

Specification

most-agreed values

Context
200K
Max output
32K
Released
2026-06-17
Knowledge cutoff
Retires
Open weights
yes
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$1.45
Output
$4.50
Cache read
$0.145

Available from 1 host

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
neuralwatt$1.45$4.50$0.145200K32K

More from Zhipu

most-hosted first