Transcribe audio
transcribeTranscribe an audio file to text using VoiceLabs' Whisper. Give the audio exactly one way: audioUrl, a public https:// URL serving the file itself (wav, mp3, m4a, aac, ogg, flac, aiff or webm; up to 10 MiB), or audioBase64, the file's real bytes (≤10 MiB of base64 text). A payload that is not real audio — under 1 KiB, a bare header, or made-up placeholder bytes — is refused without being transcribed. If the user has an audio file you cannot read the bytes of, ask them for a public link, or to upload it on the Captures page of the VoiceLabs studio (https://app.voicelabs.now/studio). Returns the transcript and the created capture id.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| audioUrl | No | A public https:// URL that serves the audio file itself (wav, mp3, m4a, aac, ogg, flac, aiff or webm; no redirects, up to 10 MiB). Give exactly one of audioBase64 or audioUrl. | |
| language | No | ISO language code (auto-detected when omitted). | |
| audioBase64 | No | The audio file's real bytes as base64 (≤10 MiB of base64 text; at least 1 KiB of audio). Only use this when you actually hold the file's bytes — a made-up or placeholder payload is refused. Give exactly one of audioBase64 or audioUrl. |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
| language | Yes | ||
| captureId | Yes | ||
| durationMs | Yes | ||
| transcript | Yes |