Skip to main content
Glama

page_snap

See a web page the way a person does: a 1280x900 screenshot (JPEG, base64) of the first screen, rendered in real headless Chrome with JavaScript executed, plus its title, full visible text and links. A failed render is never charged. $0.001 per call, USDC on Base.

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
urlYesthe http(s) page to photograph
paymentNoa signed x402 payment (the same base64 payload you would put in the PAYMENT-SIGNATURE header). Pass it here and the purchase completes inside this tool call.

Schema Changelog

Changes observed during successful MCP inspections.

  1. Added

TDQS

A4.4/5.0
Behavior4/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

With no annotations provided, the description carries the full burden of behavioral disclosure. It covers key traits: real headless Chrome with JavaScript execution (so dynamic content is rendered), a failed render is never charged (error handling), pricing per call, and payment method (USDC on Base). It also mentions the output format (JPEG, base64). This is more transparent than typical tool descriptions, though it doesn't discuss potential timeouts or large-page limitations, which are minor.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness5/5

Is the description appropriately sized, front-loaded, and free of redundancy?

The description is a single, well-structured sentence that front-loads the core purpose ('See a web page the way a person does') before specifying technical details (resolution, format, rendering). It includes essential operational details (charge, failed render policy, price, payment method) without redundancy. Every clause earns its place, making it informative yet concise.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness4/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

The tool involves significant complexity (rendering, payment, output composition) and has no output schema, yet the description covers the main outputs (screenshot, title, text, links), the rendering behavior, and the payment mechanism. It doesn't detail the exact structure of the return payload, but the list of included elements is sufficient for an agent to understand what it will receive. Given the many siblings, a more explicit differentiation would push this to a 5, but current content is adequate.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters4/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema coverage is 100%, so parameters are documented in the schema. The description adds meaningful context: it clarifies 'url' as 'the http(s) page to photograph', reinforcing the visual metaphor, and elaborates on 'payment' by explaining it's a signed x402 payment and that passing it completes the purchase within this call. This goes beyond the schema's description and aids the agent in correctly using both parameters.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

The description clearly states the tool's function: capturing a 1280x900 screenshot of a webpage rendered in headless Chrome with JavaScript, plus extracting title, visible text, and links. This specific verb-resource combination ('See a web page the way a person does') distinguishes it from siblings like page_extract, which likely focuses on text extraction. The mention of rendering and output components leaves no ambiguity about the tool's purpose.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines4/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

The description implies the appropriate use case: when a visual screenshot and full page content are needed, as opposed to text-only extraction or site-wide crawling. It does not explicitly name alternatives or provide when-not-to-use guidance, but the clear specification of output (screenshot + text + links) gives context. Slight gap in not addressing when to prefer this over a cheaper text-only tool, but the purpose is clear enough for routing.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.

Try in Browser

Glama MCP Gateway

Add one secure layer between your agents and this server.

Resources