modelbenchmark.io

GLM-4.5-Flash

Zhipu · glm-4-5-flash

Compare

Efficient GLM model for fast reasoning, coding, and agent workflows

Specification

zhipuai listing

Context
131K
Max output
98K
Released
2025-07-28
Knowledge cutoff
2025-04
Retires
Open weights
no
Input
text
Output
text

Price

US dollars per million tokens · zhipuai list price

Input
free
Output
free
Cache read
free
Cache write
free

Available from 5 hosts

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
zhipuaivendorfreefreefreefree131K98K
empiriolabsfreefree200K98K
unorouter · freefreefree131K98K
zaifreefreefreefree131K98K
zai · zaifreefree128K32K

More from Zhipu

most-hosted first