modelbenchmark.io

GLM 4.7 Flash Heretic

olafangensan-glm-4-7-flash-heretic

Compare

Efficient GLM model for fast reasoning, coding, and agent workflows

Specification

most-agreed values

Context
200K
Max output
24K
Released
2026-02-04
Knowledge cutoff
Retires
Open weights
yes
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$0.07
Output
$0.40
Cache read
$0.035

Available from 1 host

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
venice$0.07$0.40$0.035200K24K