create_text_to_speech_generation
Convert text into spoken audio using a chosen voice and model, and deliver the generated speech to workspace webhooks.
Instructions
Create Speech Generation
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| text | No | The text to synthesize into speech. | |
| voice | No | The ID of the voice to speak with. | |
| webhook | No | ||
| model_id | No | The model to use for the generation. | |
| language_code | No | ||
| output_format | No | The audio encoding of the output, as `codec_sampleRateHz_bitrateKbps`. `mp3_44100_192` requires the Creator tier or above. | |
| voice_settings | No | Overrides for the voice's saved settings, applied to one generation. | |
| pronunciation_dictionary_locators | No | Pronunciation dictionaries to apply to the text, in order of precedence. Up to 3. |