Skip to main content
Glama

apply_monophonic_audio_tuning

Plan or apply exact whole-clip pitch tuning for a monophonic Ableton Live audio clip by analyzing local source audio and verifying native pitch readback.

Instructions

Plan or apply an exact Live clip coarse/fine pitch offset from full-source, stable monophonic analysis of a mono or stereo local source at most 60 seconds long. Every stereo channel must agree; confirmation binds a SHA-256 source hash and current clip state, then verifies native pitch readback. Whole-clip tuning only: not note-by-note vocal correction, rendered audio analysis, or audible validation.

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
clipIdYesStable clip ID: track-N:clip-M from list_clips, or track-N:arrangement-clip-M from list_arrangement_clips for tools that support Arrangement clips.
dryRunNoOmit or true to return a plan; false requires a valid confirmationToken.
trackIdYesStable track ID returned by list_tracks.
planHashNoHash returned by the matching dry run.
targetMidiNoteYesExplicit equal-tempered target note at A4=440 Hz.
confirmationTokenNoShort-lived, single-use token returned by the matching dry run.
expectedStateVersionYesExact stateVersion observed immediately before planning.

Schema Changelog

Changes observed during successful MCP inspections.

  1. Addedv0.2.0

TDQS

A4.3/5.0
Behavior5/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Annotations only declare safety/idempotency flags; the description goes well beyond, disclosing the two-phase plan-then-confirm flow, SHA-256 source hashing, binding to current clip state, native pitch readback verification, the 60-second length cap, and the stereo-channel-agreement requirement. That is rich behavioral context an agent needs before invoking a mutation.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness4/5

Is the description appropriately sized, front-loaded, and free of redundancy?

Three dense sentences with no filler: purpose+constraints first, mechanism/verification second, exclusions last. Front-loaded and every clause carries information, though the middle sentence is packed tightly enough to slow reading.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness4/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

For a no-output-schema, 7-parameter mutation tool, the description covers the preconditions, safety flow, and scope limits that an agent needs. It does not describe the plan/result shape, but with no output schema that is only a minor residual gap.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is 100%, so every parameter (clipId, dryRun, planHash, confirmationToken, targetMidiNote, expectedStateVersion, trackId) is already self-documented. The description hints at the plan/confirm relationship but adds no per-parameter syntax or format detail beyond the schema, so the baseline 3 is appropriate.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

States a specific verb (plan/apply) and resource (Live clip coarse/fine pitch offset) with a clear scope qualifier: full-source monophonic analysis of a local source. An agent can tell it apart from MIDI-oriented siblings like apply_midi_transposition or MIDI plan_* tools without opening a schema.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines4/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

Clearly delimits when the tool applies (whole-clip tuning of a stable monophonic source up to 60s) and explicitly rules out adjacent cases: note-by-note vocal correction, rendered audio analysis, and audible validation. It lacks an explicit named-alternative pointer, but the exclusions do the routing work.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Deploy Server

Other Tools