modelbenchmark.io

GLM 4 32B 0414

Zhipu · glm-4-32b

Compare

Flagship GLM model for hybrid reasoning, coding, and agentic engineering

Specification

most-agreed values

Context
128K
Max output
66K
Released
2024-01-01
Knowledge cutoff
Retires
Open weights
yes
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.20
Output
$0.20
Cache read
$0.10

Available from 2 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
nano-gpt · thudm$0.20$0.20$0.10128K66K
novita$0.55$1.6632K32K

More from Zhipu

most-hosted first