Skip to main content
Glama

generate_talking_short

Start a paid Talking Short. Uses this month’s generation allowance. Poll get_talking_short until it finishes. captionMode defaults to current. phrase is MCP-ungated. The Phrase chip is operator-email only in the workspace. hook burns Current karaoke. Optional hookTemplateId seeds the brief from Opening hooks (script opening, not caption style) and does not change consume or refund.

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
urlNoProduct page URL. Must match the quote when quoteId is set.
briefNoProduct brief. Must match the quote when quoteId is set.
voiceNoNarration voice id from list_audio_lab_voices.
scriptNoTalking-head script used for this generate.
actorIdNoCatalog creator id from list_talking_short_actors.
angleIdNoAngle id from research_talking_short.
quoteIdNoQuote id from quote_talking_short. Optional; generate quotes first when omitted.
hookBlanksNoFill [bracket] keys from the Opening hooks seed. Example: { "the annoying thing": "soggy leftovers" }. Does not change the quote.
hookIntentNoOptional Opening hooks intent. Browse with list_hook_bank. Ignored when seeding if hookTemplateId is set.
aspectRatioNoFrame size. Talking Shorts is 9:16 vertical.
captionModeNoOff, Current karaoke (default), or Phrase overlays. phrase is MCP-ungated for any caller. The Phrase chip is operator-email only in the Talking Shorts workspace. hook remains accepted and burns Current karaoke (Talking Shorts has no hook_kinetic). English only. Does not change the quote amount.current
faceImageUrlNoOptional public face still URL when not using a catalog actorId.
hookTemplateIdNoOptional Opening hooks template (catalog id hook_bank). The workspace picker is hidden. Seeds topic or Talking Shorts brief when that field is empty. Fill [brackets] via hookBlanks or by editing the seeded text. This is a script opening, not a caption look (captionStyle). Does not change the quote, consume, or refund.
idempotencyKeyNoOptional key to retry an ambiguous Talking Short submit with the same inputs.
captionsEnabledNoLegacy on/off. Prefer captionStyle. false selects Off captions.
durationSecondsNoClip length in seconds. 15, 20, or 25.
ugcFinishEnabledNoPhone-cam finish on the exported MP4. Defaults on. Does not change the quote.

Output Schema

TableJSON Schema
NameRequiredDescriptionDefault
statusNoJob status such as pending, queued, in_progress, completed, or failed.
videoIdNoLibrary clip id for this job, when one exists.
idempotentNoTrue when this call reused an in-flight or finished job with the same idempotency key.
generationIdYesGeneration id. Poll the matching get_* tool until status is completed or failed.
creditsChargedNoUsage already consumed from this month’s generation allowance (internal units).

Schema Changelog

Changes observed during successful MCP inspections.

  1. First observed

TDQS

A4.7/5.0
Behavior4/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Annotations only declare readOnlyHint=false, destructiveHint=false, and openWorldHint=false, which are not very informative. The description compensates by disclosing important behaviors: 'Uses this month's generation allowance,' 'hook burns Current karaoke,' and 'hookTemplateId ... does not change consume or refund.' It also explains the captionMode nuances. No contradictions found.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is dense but highly efficient, packing critical facts into a few sentences. It front-loads the main action and immediate follow-up (poll), then clarifies key parameters without redundancy. Every sentence adds new information, and the structure is streamlined.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness5/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

Given the complexity (17 parameters, nested objects, enums) and the presence of an output schema, the description provides sufficient context for correct invocation. It covers the most complex behavioral aspects (captionMode, hookTemplateId, idempotency) and directs the agent to the right polling tool. The output schema handles return values, so nothing critical is missing.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters4/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

With 100% schema description coverage, the schema already documents each parameter well. However, the description adds extra context for specific parameters like hookTemplateId ('seeds the brief from Opening hooks, script opening, not caption style') and captionMode details about operator-email limitation. This goes beyond the schema, justifying a score above baseline 3.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description clearly states the tool's action: 'Start a paid Talking Short.' It specifies the resource (a paid Talking Short generation) and differentiates it from siblings like generate_topic_short, generate_ad, and the other generate_* tools. It also mentions key constraints like using the monthly allowance and polling get_talking_short for completion.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines5/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The description provides explicit guidance on when and how to use the tool: 'Poll get_talking_short until it finishes.' It also explains conditional behavior like captionMode defaults and the operator-email restriction for the Phrase chip. This is clear enough for an agent to know exactly what to do.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Try in Browser

Glama MCP Gateway

Add one secure layer between your agents and this server.

Resources