Skip to main content
Glama

Generate Voiceover

generation_voiceover_create
Idempotent

Queue a text-to-speech voiceover into the BlitzReels media library. Spends AI credits and returns a job to poll with generation_jobs_get.

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
textYesScript to read aloud (3-8000 characters).
speedNoSpeaking rate between 0.5 and 1.6.
voiceIdNoVoice ID. Call generation_options_list with kind voiceover for the catalog.pNInz6obpgDQGcFmaJgB
voiceStyleNoDelivery emotion.neutral
workspaceIdNoOptional workspace ID. Defaults to the user's default workspace. UUID string.
idempotencyKeyNoRetry key. Reuse only with identical inputs.

Output Schema

TableJSON Schema
NameRequiredDescriptionDefault
generationYes

Schema Changelog

Changes observed during successful MCP inspections.

  1. Changed2 schema fields changed
    • changedInput schema / properties / voiceId / default
      Previous value: -nullNew value: +"pNInz6obpgDQGcFmaJgB"
    • addedInput schema / properties / voiceId / enum
      Added value: +[
      +  "pNInz6obpgDQGcFmaJgB",
      +  "TX3LPaxmHKxFdv7VOQHJ",
      +  "FGY2WhTYpPnrIDTdsKH5",
      +  "IKne3meq5aSn9XLyUdCD",
      +  "cgSgspJ2msm6clMCkdW9",
      +  "bIHbv24MWmeRgasZH58o",
      +  "nPczCjzI2devNBz1zQrb",
      +  "JBFqnCBsd6RMkjVDRZzb",
      +  "SOYHLrjzK2X1ezoPC6cr",
      +  "N2lVS1w4EtoT3dr4eOWO",
      +  "S9EGwlCtMF7VXtENq79v",
      +  "VhxAIIZM8IRmnl5fyeyk",
      +  "hpp4J3VqNfWAUOO0d1Us",
      +  "EXAVITQu4vr4xnSDxMaL",
      +  "pFZP5JQG7iQjIQuC4Bku",
      +  "SAz9YHcvj6GT2YYXdXww",
      +  "Xb7hH8MSUJpSbSDYk0k2",
      +  "XrExE9yKIg1WjnnlVkGX",
      +  "onwK4e9ZLuTAKqWW03F9",
      +  "pqHfZKP75CvOlQylNhV4",
      +  "CwhRBWXzGAHq8TQ4Fs17",
      +  "cjVigY5qzO86Huf0OWal",
      +  "iP95p4xoKVk53GoZ742B"
      +]
  2. Changed1 schema field changed
    • changedInput schema / properties / voiceId / default
      Previous value: -"pNInz6obpgDQGcFmaJgB"New value: +null
  3. First observed

TDQS

A4/5.0
Behavior4/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Beyond the annotations (readOnlyHint=false, idempotentHint=true, destructiveHint=false, openWorldHint=true), the description adds two meaningful traits: it consumes AI credits and it produces an asynchronous job rather than a finished asset. The credit-spend warning is valuable operational context an agent cannot infer from structured fields.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

Two tight sentences with no waste. The core action ('queue a text-to-speech voiceover') is front-loaded, and the secondary facts (cost, async polling) follow in a logical order.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness4/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

An output schema exists, so return values need not be explained, and the description still covers the async nature and cost. The only minor gap is that it does not clarify workspace defaults or idempotency behavior, though the schema covers both.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is 100%, so every parameter (text, speed, voiceId, voiceStyle, workspaceId, idempotencyKey) is already documented, including the catalog lookup hint for voiceId. The description adds no syntax or format detail beyond the schema, which is the expected baseline when the schema does the heavy lifting.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

States a specific verb and resource ('Queue a text-to-speech voiceover') and names the destination ('BlitzReels media library'), which cleanly distinguishes it from siblings like generation_music_create and generation_sound_create. An agent can identify the operation without opening the schema.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines3/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

It gives follow-up guidance by directing the agent to poll with generation_jobs_get, which is genuinely useful. However, it never states when to choose this tool over the other generation_*_create siblings, so usage selection is only implied by the resource name.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Try in Browser

Glama MCP Gateway

Add one secure layer between your agents and this server.

Resources