Skip to main content
Glama

Create song generation

create_song_generation

Start a SoundBreak song generation with an AI artist. ALWAYS pass the same end_user_id for this person. Prefer artist_name from the user. Complimentary generations produce one song; a connected SoundBreak account defaults to two. After this returns, poll get_song_generation_status with generation_id and the same end_user_id until generation_status is complete. A song often takes about a minute; keep polling while poll_again is true. There is no in-chat widget — share listen_url when ready. If end_user_id is missing, do not invent one.

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
lyricsNoOptional lyrics to use instead of AI-generated lyrics
promptYesSong idea / creative direction only. Do not put take-count instructions here — set song_count instead.
song_countNoPass 1 only if the user explicitly asked for one song / one take / not two. Pass 2 if they explicitly asked for two. Omit for ordinary requests like 'write me a song'.
artist_nameNoArtist name from the user, e.g. "Kevin" or "Kevin Griffin". Preferred when the user names an artist.
end_user_idNoStable per-person id from the host. Required for complimentary songs and account linking. Pass the same value on every generate, status, library, and take-down call. Do not invent a new id each turn.
ai_cowriter_idNoArtist id from list_artists / get_artist_details. Optional if artist_name is provided.
instrumental_onlyNoIf true, generate instrumental music only

Schema Changelog

Changes observed during successful MCP inspections.

  1. Changed1 schema field changed
    • changedInput schema / properties / end_user_id / description
      Previous value: -"Stable per-person id for this caller. Required for complimentary gens and account linking. Dify: set Fixed from sys.user_id (string or number is fine). Pass the same value on every generate/status/library call. Do not invent a new id each turn."New value: +"Stable per-person id from the host. Required for complimentary songs and account linking. Pass the same value on every generate, status, library, and take-down call. Do not invent a new id each turn."
  2. First observed

TDQS

A4.8/5.0
Behavior5/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Annotations only cover the safety profile (readOnlyHint=false, openWorldHint=true, destructiveHint=false), so the description carries the rest and does it well: it discloses the complimentary vs. connected-account song-count default, the roughly one-minute latency, the polling contract with poll_again, and the absence of an in-chat widget. These are behaviors an agent could not infer from the schema.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness4/5

Is the description appropriately sized, front-loaded, and free of redundancy?

Built from short, imperative sentences that front-load the action, then the follow-up polling flow, then the failure guard. It is slightly repetitive on end_user_id consistency, which is stated in the description and again in the schema, but no sentence is wasted.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness5/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

For a 7-parameter generation tool with no output schema and no nested objects, the description supplies everything an agent needs: what returns (generation_id to poll), how long it takes, how to detect completion, and what to do with the result. Nothing material is left to inference.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters4/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema coverage is 100%, so the baseline is 3, but the description adds cross-parameter semantics beyond the schema: reusing the same end_user_id across every generate/status/library/take-down call, preferring artist_name over ai_cowriter_id, and the rule that the account tier (not the request) determines whether one or two songs come back. Some of this restates schema text, keeping it off a 5.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

Opens with a specific verb+resource pair ('Start a SoundBreak song generation with an AI artist'), naming the product and the actor. It is immediately distinguishable from siblings like get_song_generation_status, list_my_songs, or remove_song_preview.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines5/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

Explicitly routes to the alternative: poll get_song_generation_status with generation_id and the same end_user_id until generation_status is complete, keeping at it while poll_again is true. It also gives a pre-condition (do not invent end_user_id if missing) and a post-action (share listen_url, no in-chat widget).

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Try in Browser

Glama MCP Gateway

Add one secure layer between your agents and this server.

Resources