text_to_voice_remix
Modify an existing voice by applying described changes to its characteristics. Produces a new voice variant from a voice ID and a change description.
Instructions
Remix A Voice. Spends ElevenLabs credits.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| seed | No | ||
| text | No | ||
| loudness | No | Controls the volume level of the generated voice. -1 is quietest, 1 is loudest, 0 corresponds to roughly -24 LUFS. | |
| voice_id | Yes | Voice ID to be used, you can use https://api.elevenlabs.io/v1/voices to list all the available voices. | |
| output_format | No | Output format of the generated audio. Formatted as codec_sample_rate_bitrate. So an mp3 with 22.05kHz sample rate at 32kbs is represented as mp3_22050_32. MP3 with 192kbps bitrate requires you to be subscribed to Creator tier or above. PCM with 44.1kHz sample rate requires you to be subscribed to Pr | |
| guidance_scale | No | Controls how closely the AI follows the prompt. Lower numbers give the AI more freedom to be creative, while higher numbers force it to stick more to the prompt. High numbers can cause voice to sound artificial or robotic. We recommend to use longer, more detailed prompts at lower Guidance Scale. | |
| prompt_strength | No | ||
| stream_previews | No | Determines whether the Text to Voice previews should be included in the response. If true, only the generated IDs will be returned which can then be streamed via the /v1/text-to-voice/:generated_voice_id/stream endpoint. | |
| voice_description | Yes | Description of the changes to make to the voice. | |
| auto_generate_text | No | Whether to automatically generate a text suitable for the voice description. | |
| remixing_session_id | No | ||
| remixing_session_iteration_id | No |