read_web_page
Extract clean, readable Markdown and metadata from any public webpage for LLM ingestion, stripping ads, popups, and navigational clutter. Returns clean markdown, title, description, character count, and estimated tokens. Requires x402 micropayment (0.005 USDC on Base).
When to use: Ingesting articles, blog posts, documentation, or news pages into LLM context. When NOT to use: Do NOT use for raw binary files (PDF/images), authenticated pages behind a login, or single-page apps that require heavy JavaScript rendering.
Parameters:
url(string, required): Full target webpage URL (e.g. 'https://news.ycombinator.com').include_links(boolean, optional, default true): Whether to preserve markdown hyperlinks.include_images(boolean, optional, default false): Whether to preserve image markdown links.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| url | Yes | Full target webpage URL to extract markdown from. | |
| include_links | No | Preserve markdown hyperlinks. | |
| include_images | No | Preserve markdown image tags. |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
| url | Yes | Target webpage URL | |
| title | No | Extracted page title | |
| length | Yes | Content length in characters | |
| content | Yes | Clean extracted Markdown content | |
| elapsed_ms | No | Extraction time in milliseconds | |
| description | No | Extracted meta description | |
| estimated_tokens | No | Estimated LLM tokens (~len/4) |