modelbenchmark.io

Grok Voice STT 1.0

xAI · grok-voice-stt-1-0

Compare

Grok Voice STT 1.0 is xAI's speech-to-text model. It supports transcription with word-level timestamps, optional speaker diarization, and multichannel audio.

Specification

most-agreed values

Context
15K
Max output
15K
Released
2026-08-04
Knowledge cutoff
—
Retires
—
Open weights
no
Input
audio
Output
text

Available from 1 host

HostIn $/MOut $/MCache rdCache wrContextOutputRetires
zenmux · x-ai————15K15K—

More from xAI

most-hosted first