modelbenchmark.io

GLM-4 Long

Zhipu · glm-4-long

Compare

Compact GPT model for low-latency assistance and high-volume workloads

Specification

most-agreed values

Context
1M
Max output
4K
Released
2024-01-01
Knowledge cutoff
Retires
Open weights
no
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.2006
Output
$0.2006
Cache read
$0.1003

Available from 1 host

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
nano-gpt$0.2006$0.2006$0.10031M4K

More from Zhipu

most-hosted first