v6

latestOpenAPI 3.0.1raw.githubusercontent.com2026-08-0182128753.1 KB
Audio

Convert Text to Audio

Available for: Chatflow, Workflow, New Agent, Chatbot, Agent, Text Generator apps.

Converts text to speech audio. Pass text to synthesize arbitrary text, or message_id to voice an existing message's answer.

post/text-to-audio

Request body

message_idstring uuid

ID of the message whose answer to voice. Takes priority over text when both are provided. Get message IDs from List Conversation Messages.

textstring

Text to synthesize into speech.

userstring

End-user identifier, defined by your app and unique within it. See End User Identity.

voicestring

Voice to use for text-to-speech. Available voices depend on the TTS provider configured for this app. Use the voice value from Get App Parameterstext_to_speech.voice for the default.

streamingboolean

Accepted for backward compatibility but has no effect. Whether the audio is streamed is determined by the configured TTS provider's output, not by this field.

Response

Returns the generated audio file. The Content-Type header is set to the audio MIME type (e.g., audio/wav, audio/mp3). When the configured TTS provider returns a stream, the audio is delivered as audio/mpeg with chunked transfer encoding; the request streaming field does not control this.