@page2ai/mcp
This server provides a single tool, page_to_markdown, that fetches any public web page and converts it into clean Markdown optimized for LLM context — running entirely locally with no external API calls.
Core capabilities:
Convert any public URL to Markdown: Provide an absolute HTTP/HTTPS URL; the server fetches and parses it locally using
linkedom.Clean, structured output: Preserves headings, code blocks (with language hints), links, and tables while automatically stripping ads, navigation bars, cookie banners, and other clutter.
Mintlify docs optimization: For Mintlify-hosted documentation (e.g., docs.anthropic.com, OpenAI, Vercel, Stripe), it tries the
URL.mdconvention for the cleanest output.Tab group handling: Discovers tab groups (e.g., Python vs. TypeScript examples) and emits each panel as a separate
### Tab: {label}section, preventing code samples from merging incorrectly.Configurable options:
timeout_ms: 1,000–60,000ms (default: 15,000ms)include_images: include/exclude image references (default:true)include_frontmatter: prepend YAML metadata (title, URL, timestamp, etc.) (default:true)
Security: Blocks private IP ranges, loopback addresses, and cloud metadata endpoints to prevent SSRF-style misuse.
Privacy-friendly: Zero telemetry, no data collection, no external network calls beyond the explicitly provided URL.
Broad client support: Works with Claude Desktop, Cursor, Windsurf, Zed, VS Code, and any MCP-compatible client.
Example use cases: Fetch documentation to answer questions, extract API references into code snippets, or compare multiple documentation pages.
Click on "Install Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@@page2ai/mcpConvert https://docs.anthropic.com/en/api/messages to clean Markdown"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
@page2ai/mcp — Web to Markdown for LLM Context
Turn any web page into clean Markdown for Claude, ChatGPT, or your own LLM.
Companion to the Page2AI Chrome extension. Shares the same @page2ai/core extraction library. Zero external API calls — runs entirely on your machine using linkedom.
Install
Claude Desktop
Add to ~/Library/Application Support/Claude/claude_desktop_config.json (macOS) or %APPDATA%/Claude/claude_desktop_config.json (Windows):
{
"mcpServers": {
"page2ai": {
"command": "npx",
"args": ["-y", "@page2ai/mcp"]
}
}
}Restart Claude Desktop.
Cursor
Add to ~/.cursor/mcp.json (global) or .cursor/mcp.json (project):
{
"mcpServers": {
"page2ai": {
"command": "npx",
"args": ["-y", "@page2ai/mcp"]
}
}
}Windsurf
Add to ~/.windsurf/mcp.json:
{
"mcpServers": {
"page2ai": {
"command": "npx",
"args": ["-y", "@page2ai/mcp"]
}
}
}Zed
Add to settings.json:
{
"context_servers": {
"page2ai": {
"command": {
"path": "npx",
"args": ["-y", "@page2ai/mcp"]
}
}
}
}VS Code (with Continue or GitHub Copilot Chat)
Refer to your MCP-compatible extension's documentation. The command is npx -y @page2ai/mcp.
Related MCP server: Webustler
Tools
Tool | Description | Read-only | Example |
| Fetch a web page URL and convert to clean Markdown | ✅ |
|
Example prompts
1. Fetch documentation and answer questions:
"Use page_to_markdown to fetch https://docs.anthropic.com/en/docs/build-with-claude/extended-thinking and summarize the three main use cases for extended thinking."
2. Extract API reference into a code snippet:
"Fetch https://ai.google.dev/gemini-api/docs/thinking with page_to_markdown, then generate a Python code sample using the thinking budget parameter."
3. Compare two documentation pages:
"Fetch both https://docs.anthropic.com/en/docs/prompt-engineering and https://platform.openai.com/docs/guides/prompt-engineering with page_to_markdown, then summarize the differences in approach."
Configuration
None required in v0.1. All extraction options use sensible defaults.
Future versions will support options via tool arguments:
include_images(boolean, defaultfalse)include_frontmatter(boolean, defaulttrue)profile(string, one ofauto | docs | marketing | research | dashboard | wordpress-marketing, defaultauto)
Privacy
@page2ai/mcp collects no data, sends no telemetry, and makes no external network calls beyond the URLs you explicitly provide. See PRIVACY.md for details.
Protocol revisions
Since 0.2.0 this server answers both MCP protocol revisions on the same stdio connection:
Revision | How a client opens the connection | Served |
|
| yes |
| the | yes |
serveStdio picks the era from the opening message and pins one server instance for
the life of the connection, so clients on older SDKs are unaffected. Note that a
message carrying no version claim is treated by the specification as a 2025-era
opening; server/discover without that _meta field is therefore answered with
-32601 rather than being upgraded.
Verify against a build with the raw JSON-RPC frames, without a client library:
printf '%s\n' '{"jsonrpc":"2.0","id":1,"method":"server/discover","params":{"_meta":{"io.modelcontextprotocol/protocolVersion":"2026-07-28"}}}' | node dist/index.jsKnown advisories
None. Moving to the v2 SDK removed 90 transitive packages (117 production
dependencies down to 26) and with them the @hono/node-server advisory that
0.1.x carried through the monolithic @modelcontextprotocol/sdk; that package
pulled the HTTP and OAuth server stack even for a stdio-only server. npm audit
is clean as of 0.2.0.
Development
git clone https://github.com/igorsaevets/page2ai-mcp
cd page2ai-mcp
npm install
npm run build
node dist/index.js # runs stdio MCP serverTest with MCP Inspector:
npx @modelcontextprotocol/inspector node dist/index.jsSupport
Pick the channel by what you have, not by what is quickest to type:
You have | Use |
A bug, a page that extracts badly, a feature idea | GitHub issues — public and searchable, so the fix helps the next person too |
A security vulnerability | Private vulnerability reporting, see SECURITY.md. Do not open a public issue |
A usage question | Docs first, then an issue |
About the author
Written and maintained by Igor Saevets — AI expert and founder of Page2AI. Full bio: igorsaevets.github.io/page2ai-docs/about/.
These are identity and collaboration links, not the support queue. A bug reported in a DM is a bug nobody else can find later, so anything you want fixed belongs in the table above.
LinkedIn: linkedin.com/in/igorsaevets
Facebook: facebook.com/igorsaevets
GitHub: github.com/igorsaevets
ORCID: 0009-0006-8636-1377
Email: igorsaevets@gmail.com
Build provenance
An MCP server runs with the privileges of whatever launched it, so where the tarball came from is a security question, not a formality. Releases from v0.1.2 onward are published from GitHub Actions with npm provenance: each version carries a Sigstore attestation naming the commit and workflow run that produced it, recorded in the public Rekor transparency ledger and shown as a badge on npmjs.com.
Check it before you trust it:
npm audit signatures
npm view @page2ai/mcp --json | jq '.dist.attestations'License
MIT — see LICENSE. Copyright © 2026 Igor Saevets.
Related
Page2AI Chrome extension — https://github.com/igorsaevets/page2ai-extension (same extraction core, distributed as a browser extension for humans)
@page2ai/core — https://npmjs.com/package/@page2ai/core (the shared extraction library)
page2ai-mcp (unscoped) — https://npmjs.com/package/page2ai-mcp (thin wrapper around this package, published so
npx -y page2ai-mcpworks without a scope prefix)Software Heritage archive — SWHID
swh:1:snp:05123c51ef9e7c0aeb06f42b1263c07a8d26999a
Maintenance
Tools
Related MCP Servers
- AlicenseAqualityDmaintenanceConverts URLs and raw HTML to clean Markdown, enabling AI assistants to read web pages for summarization, analysis, or ingestion.2171MIT
- Alicense-qualityCmaintenanceEnables clean, LLM-ready markdown extraction from any URL with automatic anti-bot bypass.3MIT
- Alicense-qualityDmaintenanceConverts any webpage into clean, LLM-ready Markdown, removing noise and supporting JavaScript rendering.MIT
- AlicenseAqualityBmaintenanceConverts any URL to clean, LLM-ready Markdown by removing ads, navigation, and other clutter, using Mozilla Readability and Turndown.157MIT
Related MCP Connectors
Converts any URL to clean, LLM-ready Markdown using real Chrome browsers
Web scraping for AI agents. Converts URLs to clean, LLM-ready Markdown with anti-bot bypass.
Fetch any URL and get clean Markdown. Web scraping for AI agents.
Latest Blog Posts
- Who's Calling? MCP Hosts Are an Identity Blind Spot (And the Spec Knows It)By Om-Shree-0709 on .mcpAgent IdentityOAuth 2.1
- Your AI Chatbot Just Exposed Your CEO's Salary to an InternBy Om-Shree-0709 on .Agent IdentityMCP SecurityOAuth Delegation
- Why MCP Servers Need Execution Sandboxing (And Why Your Current Stack Isn't Enough)By Om-Shree-0709 on .Agentic AiPrompt InjectionWebAssembly
MCP directory API
We provide all the information about MCP servers via our MCP API.
curl -X GET 'https://glama.ai/api/mcp/v1/servers/igorsaevets/page2ai-mcp'
If you have feedback or need assistance with the MCP directory API, please join our Discord server