Actions7
Generate Speech from Prompt
AI-generatedSummary
Generates spoken audio from input text using the Picsart GenAI text-to-speech API.
Inputs
- text2SpeechText (required) — The text to speak (up to 5000 characters).
- text2SpeechLanguage — The language to speak (defaults to 'en').
- text2SpeechModel — AI model for text-to-speech (defaults to OpenAI tts-1).
- text2SpeechVoice — Optional voice name for synthesis; defaults to provider-specific defaults.
Output shape
One item per input item containing a binary audio file and metadata.
Audio is polled until ready (max 600 attempts). Returns an MP3 file named 'generated-speech.mp3'.
Examples
Example 1: Convert a static text string to audio in the English language.
Set text2SpeechText to 'Hello world' and leave other parameters at defaults.
Example 2: Generate speech in French with a specific provider model.
Set text2SpeechText to 'Bonjour', text2SpeechLanguage to 'fr', and text2SpeechModel to ElevenLabs.