modelbenchmark.io

Hermes Low

hermes-low

Compare

Compact GPT model for low-latency assistance and high-volume workloads

Specification

most-agreed values

Context
1M
Max output
131K
Released
2026-05-11
Knowledge cutoff
Retires
Open weights
no
Input
text
Output
text

Price

US dollars per million tokens · most-agreed

Input
$1.00
Output
$3.20
Cache read
$0.20
Cache write
$0.0833

Available from 1 host

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
nano-gpt$1.00$3.20$0.20$0.08331M131K