Type identifier:
ai:text-to-speechCategory: AI Operations
The AI - Text to Speech node generates realistic spoken audio from text input using OpenAI's text-to-speech models. It supports multiple voice options and quality levels, returning an audio file that can be used in Experiences or stored.
Handle | Type | Description |
|---|---|---|
|
| The text to convert to speech. |
|
| (Dynamic) The TTS model to use. |
|
| (Dynamic) The voice to use. |
Note:
modelandvoicehandles appear when set to dynamic mode.
Handle | Type | Description |
|---|---|---|
|
| The generated audio file. |
Option | Type | Default | Description |
|---|---|---|---|
Text |
| dynamic | The text content to synthesize. |
Option | Type | Default | Description |
|---|---|---|---|
Model | literal or dynamic |
| The TTS model to use. |
See AI - Model Reference — Text to Speech Models for the canonical model list.
Available models:
Model | Quality | Speed | Use Case |
|---|---|---|---|
| Standard | Fast | Real-time applications |
| High Definition | Slower | High-quality output |
Option | Type | Default | Description |
|---|---|---|---|
Voice | literal or dynamic |
| The voice character to use. |
Available voices:
Voice | Character |
|---|---|
| Neutral, balanced |
| Warm, natural |
| Expressive, British |
| Deep, authoritative |
| Friendly, energetic |
| Clear, pleasant |
Generated audio is typically in MP3 format.
OpenAI TTS has a character limit per request. Long text may need to be chunked.
Configuration:
openai:tts-1openai:novaInput (text):
"Welcome to our application. Let me guide you through the features."Output: MP3 audio file with spoken content.
Configuration:
openai:tts-1-hdopenai:onyxUse case: Generate professional-quality audio for podcasts or presentations.
Configuration:
openai:tts-1Use case: Allow user to select their preferred voice at runtime.