Clone a voice from a recording
clone_voice_profileCreate a cloned voice on the user's VoiceLabs account from a recording the caller provides: a name, the language, the exact text spoken in the recording, and the audio either inline as base64 (audioBase64) or as an https:// URL (audioUrl). The recording must be 2–30 seconds of clear speech. Returns the new profile's id, name, voice type and engine; pass the id to speak. Not idempotent: a second call with the same name is refused, so check list_voice_profiles first. Only clone a voice whose owner has agreed to it.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| name | Yes | The name for the new cloned voice. Must not already exist on the account — call list_voice_profiles first if unsure; a duplicate is refused with INVALID_REQUEST. | |
| audioUrl | No | An https:// URL the recording can be downloaded from (public host, no redirects, up to 10 MiB). Give exactly one of audioBase64 or audioUrl. | |
| language | No | ISO language code of the recording and the voice. Defaults to "en". | |
| audioBase64 | No | The recording inline as base64 (WAV, or any of mp3/m4a/ogg/flac/aac/webm/opus; up to 10 MiB decoded). 2–30 seconds of clear speech. Give exactly one of audioBase64 or audioUrl. | |
| referenceText | Yes | The exact words spoken in the recording (up to 1000 characters). The engine aligns the clone against this transcript, so it must match what is said. |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
| id | Yes | ||
| name | Yes | ||
| engine | Yes | ||
| language | Yes | ||
| sampleId | Yes | ||
| voiceType | Yes |