Skip to main content
Glama

youtube_get_transcript

Fetch the transcript and metadata for a single YouTube video from a URL or video ID. Results are cached locally for efficient reuse.

Instructions

Fetch the transcript and metadata for a single YouTube video. Accepts a URL or video ID. Results are cached locally. Fetch only the videos most relevant to the user's question — avoid bulk fetching. The transcript field is an array of segments, each with text, start (seconds), and duration (seconds).

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
video_urlYes

Output Schema

TableJSON Schema
NameRequiredDescriptionDefault
resultYes
Behavior3/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

With no annotations provided, the description carries the full burden of behavioral disclosure. It mentions local caching and the segment structure of transcripts, adding useful context. However, it does not state whether the operation is read-only, potential errors, or authentication requirements.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is three sentences, each providing unique and necessary information: purpose, usage constraint, and result format. It is front-loaded and free of filler.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness4/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

For a tool with only one parameter and an output schema, the description covers the core purpose, usage context, parameter format, and even adds structural details about the transcript. It could mention error handling or rate limits, but these are not critical for a straightforward fetch operation.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters4/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

The schema only defines video_url as a string, but the description clarifies that it accepts a URL or video ID, which is essential for correct invocation. With 0% schema description coverage, this added meaning significantly helps the agent.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description clearly states the tool fetches transcript and metadata for a single YouTube video, using a URL or video ID. It distinguishes itself from the sibling tool youtube_search by specifying the exact resource (transcript) and action (fetch).

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines4/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The description provides explicit guidance to 'fetch only the videos most relevant to the user's question — avoid bulk fetching,' which helps the agent decide when to use this tool. It does not explicitly contrast with youtube_search, but the context is clear enough.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Install Server

Other Tools

Latest Blog Posts

MCP directory API

We provide all the information about MCP servers via our MCP API.

curl -X GET 'https://glama.ai/api/mcp/v1/servers/BlockBenny/tubemcp'

If you have feedback or need assistance with the MCP directory API, please join our Discord server