Skip to main content
Glama

sitecheck

read_url

Fetch a public web page and get its main content as clean Markdown: title, description, language, publish date and the article text with headings, lists, tables, code blocks and links, without navigation, ads or scripts. Built for LLM agents that need to read a URL. Respects robots.txt; static HTML only (no JavaScript rendering). Not charged if the page cannot be read. 5 free calls per day across all tools; pass key to use a prepaid balance (0.005 USDC per call).

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
keyNoOptional API key from create_key
urlYesAbsolute http(s) URL of the page to read
linksNoSet to 0 to drop link targets and keep only the link text
max_charsNoMaximum characters of content to return (default 20000, maximum 100000)

Schema Changelog

Changes observed during successful MCP inspections.

  1. First observed

TDQS

A4.4/5.0
Behavior5/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

With no annotations, the description carries the full burden and largely does so: robots.txt compliance, no JS rendering, stripping of nav/ads/scripts, 'not charged if the page cannot be read', and rate/pricing limits (5 free calls/day, 0.005 USDC per call via prepaid key). These are non-obvious behavioral traits an agent must know before calling.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness4/5

Is the description appropriately sized, front-loaded, and free of redundancy?

Purpose and output shape are front-loaded, followed by constraints and then billing. Dense but every sentence carries information; the billing sentence is slightly packed but still earns its place.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness5/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

No output schema and no annotations, yet the description covers what the return payload contains, what is stripped, rendering limitations, compliance behavior, and cost model. Nothing an agent needs to call it correctly is missing.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters3/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema coverage is 100%, so the schema already documents url, links, max_chars and key. The description adds only pricing/billing context for key ('pass key to use a prepaid balance') and does not explain links or max_chars behavior. Baseline 3 is appropriate when the schema does the heavy lifting.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

States a specific verb and resource ('Fetch a public web page') and enumerates exactly what is returned: title, description, language, publish date, and article text with headings, lists, tables, code blocks and links. This is far more specific than any sibling (balance, check_domain, create_key), so an agent can route to it unambiguously.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines4/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

Gives clear intended context ('Built for LLM agents that need to read a URL') and meaningful exclusions: static HTML only, no JavaScript rendering, and robots.txt is respected. It does not explicitly compare against a sibling alternative, but the siblings are largely unrelated, so the boundary is inferable.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Try in Browser

Glama MCP Gateway

Add one secure layer between your agents and this server.

Resources