Generate music
generate_musicTurn text prompts or audio files into music, covers, remixes, stems, completions, or analysis with the Audial model.
Instructions
Generate music with the Audial music model.
Text to music, covers, remixes, stem extraction, completion, or analysis (understand).
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| bpm | No | Tempo hint in beats per minute. | |
| seed | No | Seed for reproducible output. | |
| lyrics | No | Lyrics with optional section tags like [Verse] and [Chorus]. Omit for instrumental. | |
| prompt | Yes | Style description: genre, mood, instruments, vocal character. Tempo and key words in the prompt steer the model more than the numeric bpm/key fields, which are hints, not constraints. | |
| key_scale | No | Key hint, e.g. 'G major'. | |
| task_type | No | text2music (default), cover, remix, extract, lego, complete, or understand. | text2music |
| batch_size | No | Number of variations, 1-8. | |
| track_name | No | Track to extract/replace for extract/lego: vocals, drums, bass, guitar, piano, strings, synth, other. | |
| source_file | No | Local audio file; required for remix, extract, lego, complete, understand. | |
| audio_format | No | mp3, wav, flac, opus or aac. | mp3 |
| instrumental | No | Generate without vocals. | |
| audio_duration | No | Length in seconds (10-600). | |
| reference_file | No | Local audio file; required for cover. | |
| repainting_end | No | Remix end time in seconds (-1 = end). | |
| time_signature | No | e.g. '4/4'. | |
| vocal_language | No | Language code for vocals, e.g. 'en'. | en |
| negative_prompt | No | What to avoid, comma-separated. | |
| repainting_start | No | Remix start time in seconds. | |
| audio_cover_strength | No | 0-1 fidelity to the original for cover/remix. |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
| tool | Yes | ||
| files | Yes | ||
| summary | Yes | ||
| metadata | Yes | ||
| output_dir | Yes | ||
| execution_id | Yes |