One of the available TTS models: tts-1, tts-1-hd, gpt-4o-mini-tts, or gpt-4o-mini-tts-2025-12-15.
The text to generate audio for. The maximum length is 4096 characters.
Control the voice of your generated audio with additional instructions. Does not work with tts-1 or tts-1-hd.
A built-in voice name or a custom voice reference.
The format to audio in. Supported formats are mp3, opus, aac, flac, wav, and pcm.
The speed of the generated audio. Select a value from 0.25 to 4.0. 1.0 is the default.
The format to stream the audio in. Supported formats are sse and audio. sse is not supported for tts-1 or tts-1-hd.
{ "voice": { "id": "voice_1234" } }
OK