webpage-readability-mcp
Click on "Deploy Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@webpage-readability-mcpextract the main article from https://example.com/blog/post"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
webpage-readability-mcp
Read-only MCP server that fetches a single URL and extracts its main content -- the article/post text, stripped of navigation, ads, sidebars, and boilerplate -- using trafilatura. This is the "readability mode" pattern (Firefox Reader View, Safari Reader, the original Arc90 Readability bookmarklet) as an MCP tool.
Tools
Tool | Description |
| Fetches |
Related MCP server: Scrapi MCP Server
Setup
cd webpage-readability-mcp
python -m venv .venv
.venv/Scripts/activate # Windows; use `source .venv/bin/activate` on macOS/Linux
pip install -r requirements-dev.txtRun the tests
python -m pytest -qRegister with Claude Code
claude mcp add webpage-readability -- E:/repo/unt/mcp/webpage-readability-mcp/.venv/Scripts/python.exe E:/repo/unt/mcp/webpage-readability-mcp/server.pyOr add it manually to your MCP config (e.g. .claude/settings.json or the global Claude Code MCP config):
{
"mcpServers": {
"webpage-readability": {
"command": "E:/repo/unt/mcp/webpage-readability-mcp/.venv/Scripts/python.exe",
"args": ["E:/repo/unt/mcp/webpage-readability-mcp/server.py"]
}
}
}Limitations
No JavaScript rendering -- this is a plain HTTP GET + HTML parse, not a headless browser. A JS-heavy SPA whose content is injected client-side will extract poorly or empty.
No authentication, cookies, or paywall handling -- every fetch is anonymous; a paywalled or login-gated page returns whatever the anonymous response contains.
Single-fetch-per-call only -- exactly one HTTP request per call, to exactly the URL given. No pagination-following, no link-following, no caching.
No PDF or other non-HTML content extraction -- a URL that doesn't resolve to an HTML/XHTML content-type returns an error rather than being silently mis-parsed.
This server cannot be deployed
Maintenance
Related MCP Connectors
MCP server (stdio): fetch web pages as clean readable markdown via the AgentForge API
Screenshot, PDF, OG-image, and page extraction (markdown/JSON) over MCP. Bearer key or x402.
Hosted RoboWrite MCP server: recover briefs and drafts, run the reviewed URL-to-Markdown workflow.
Document-to-Markdown MCP server — convert PDF, Office and HTML into LLM-ready Markdown.
Related MCP Servers
- AlicenseNot gradedqualityFmaintenanceAn MCP server that extracts clean Markdown or HTML content from web pages by stripping away ads, navigation, and clutter. It offers tools to process URLs or raw HTML, returning structured metadata alongside the main article content.2MIT
- AlicenseAqualityCmaintenanceMCP server that converts URLs to clean Markdown/Text for LLM agents.2553 npm5MIT
- AlicenseNot gradedqualityDmaintenanceAn MCP server for intelligent web content extraction from JavaScript-heavy sites using single-file and trafilatura. It enables AI agents to fetch, render, and paginate through clean article content and metadata.20MIT
- FlicenseNot gradedqualityDmaintenanceThis MCP server enables clean web content extraction from URLs or HTML using Trafilatura, supporting multiple output formats and configurable extraction options.-