Grok STTDisabled since Jul 26, 2026
The catalogue was narrowed to text- and embedding-output models. This model's output IS text, but its input is audio — it is a speech-to-text SKU, named explicitly in the same decision, so it goes out of service with the image / video / TTS models rather than being kept on the technicality of its output modality.
xAI's Grok STT is a speech-to-text transcription model available as a partner-hosted SKU on Cloudflare Workers AI. Converts audio input to text.
Share cards2 images

