forced_alignment
Aligns text with an audio file to map spoken words to timestamps; consumes ElevenLabs credits and supports common audio formats up to 1GB.
Instructions
Create Forced Alignment Spends ElevenLabs credits.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| text | Yes | The text to align with the audio. The input text can be in any format, however diarization is not supported at this time. | |
| file_path | No | The file to align. All major audio formats are supported. The file size must be less than 1GB. Local path. Required for this call. | |
| file_base64 | No | Base64 contents for "file". Use this when the server cannot read your local disk. | |
| file_filename | No | Filename to send for "file". Some endpoints infer the audio format from it. |