Skip to main content
Glama
MarekCziba

Website to Markdown MCP

by MarekCziba

extract_metadata

Extracts page metadata such as title, description, author, and Open Graph tags from an absolute HTTP(S) URL for use by AI agents.

Instructions

Extract page metadata (title, description, author, Open Graph tags).

Args: url: Absolute http(s) URL of the page to inspect.

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
urlYes

Output Schema

TableJSON Schema
NameRequiredDescriptionDefault

No arguments

Schema Changelog

Changes observed during successful MCP inspections.

  1. First observedv0.1.0

TDQS

B3.4/5.0
Behavior2/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

No annotations are provided, so the description carries the full behavioral burden. It does not say whether the tool fetches the URL over the network, how redirects or unreachable/timeout URLs are handled, whether auth or robots restrictions apply, or whether a head-only request is made. 'Inspect' is the only hint about mechanism.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness4/5

Is the description appropriately sized, front-loaded, and free of redundancy?

One front-loaded sentence naming the tool's output, followed by a compact arg note with no filler. The 'Args:' block is slightly redundant with a single obvious parameter but is not wasteful.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness4/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

Output is covered by an output schema, so return values need no explanation, and a single required parameter keeps the surface small. The remaining gap is the absent behavioral detail around network access and failure handling, which matters for a URL-fetching tool but is minor given overall simplicity.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters4/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema coverage is 0% and the schema only types url as a bare string, so the description's 'Absolute http(s) URL of the page to inspect' meaningfully constrains the expected format. For a one-parameter tool this is adequate compensation, though it adds no detail on encoding or query-string handling.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose4/5

Does the description clearly state what the tool does and how it differs from similar tools?

States a specific verb ('Extract') and resource ('page metadata') and enumerates the extracted fields (title, description, author, Open Graph tags). It is distinguishable from fetch_url/fetch_urls, which retrieve content rather than parsed metadata, though it never names those siblings explicitly.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines3/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

Usage is only implied by contrast with the sibling fetch tools: an agent can infer this is for metadata rather than full page content. There is no explicit when-to-use, when-not-to-use, or alternative routing statement.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Deploy Server

Other Tools