Grok Voice TTS 1.0
xAI · 1 configuration
Intelligence—
Input / 1M—
Output / 1M—
Context15K
Tariff: ZenMux · Source: models-dev · USD per million tokens
Thinking configuration
What does more thinking buy you?
Compare this model’s measured thinking configurations.
◌
No pair of measurements available
Change the filters or pick another benchmark. Missing values are not estimated.
Selected: thinking default · Click a point to switch configuration
Convert text into spoken audio with a single API call. The API supports a rich set of expressive voices, inline speech tags for fine-grained delivery control, and output formats from high-fidelity MP3 to telephony-optimized μ-law.
This model is in the catalog but has no imported benchmarks. Capabilities are not estimated.
Specifications from the catalog
Inputtext
Outputaudio
Tool callingNo
Structured outputNot available
SpeedNot available
First tokenNot available
Maximum output15K
Catalog listing date2026-07-31