Website Content Crawler MCP Server
Website Content Crawler MCP Server
Wrapper MCP que ejecuta realmente el actor apify/website-content-crawler mediante apify-client y devuelve sus páginas como JSON. Cada llamada devuelve contenido rastreado o un error estructurado — nunca informa de un éxito falso.
Cuánto cuesta
Las ejecuciones se facturan a tu cuenta de Apify según la tarifa indicada para apify/website-content-crawler — este servidor no añade nada por encima. Sin token, sin cargo: las llamadas devuelven missing_token antes de que comience ninguna ejecución.
Demostración en vivo de la salida (misma lógica de rastreo, renderizada): https://website-content-crawler.vercel.app
Related MCP server: Crawl4AI MCP Server
Configuración
npm install
export APIFY_TOKEN=apify_api_... # https://console.apify.com/settings/integrations
npm start # stdio MCP serverHerramienta: crawl_website
Entrada:
campo | type | por defecto | notas |
| string | obligatorio | URL http/https para rastrear |
| number | 10 | con un tope de 50 |
| string |
| o |
Formato de salida (la misma estructura en cada llamada):
{
"status": "ok",
"run": { "id": "<apify run id>", "status": "SUCCEEDED" },
"page_count": 3,
"total_in_dataset": 3,
"pages": [{ "url": "...", "title": "...", "text": "...(≤5000 chars)" }]
}Códigos de error: invalid_url, missing_token, apify_auth_failed, actor_run_failed, run_not_succeeded, dataset_fetch_failed.
Prueba de humo
printf '%s\n' \
'{"jsonrpc":"2.0","id":1,"method":"initialize","params":{"protocolVersion":"2024-11-05","capabilities":{},"clientInfo":{"name":"t","version":"0"}}}' \
'{"jsonrpc":"2.0","method":"notifications/initialized"}' \
'{"jsonrpc":"2.0","id":2,"method":"tools/list"}' \
'{"jsonrpc":"2.0","id":3,"method":"tools/call","params":{"name":"crawl_website","arguments":{"url":"https://example.com"}}}' \
| node index.jsSin APIFY_TOKEN definido, la solicitud con id 3 debe devolver
{"status":"error","error":{"code":"missing_token",...}} — prueba de que el manejador
alcanza el límite real de Apify en lugar de inventarse un resultado.
Maintenance
Resources
Unclaimed servers have limited discoverability.
Looking for Admin?
If you are the server author, to access and configure the admin panel.
Tools
Related MCP Servers
- AlicenseNot gradedqualityDmaintenanceEnables web scraping and document processing with JavaScript execution, anti-detection measures, batch processing, and structured data extraction. Supports multiple formats including markdown, HTML, screenshots, and handles PDFs with OCR capabilities.3MIT
- FlicenseNot gradedqualityDmaintenanceEnables advanced web crawling and content extraction with JavaScript support, AI-powered analysis, PDF/Office document processing, YouTube transcript extraction, Google search integration, and multi-format data export capabilities.2
- FlicenseNot gradedqualityDmaintenanceProvides web crawling and browser automation capabilities with support for multiple content formats (HTML, JSON, PDF, screenshots, Markdown), page content extraction, console message monitoring, and network request tracking.
- AlicenseAqualityAmaintenanceEnables web scraping, structured data extraction, and screenshot capture with automatic anti-bot bypass, supporting JavaScript rendering, proxy rotation, and tiered pricing.251871MIT
Related MCP Connectors
Scrape, crawl, map and extract the web. Pay per call in USDC, no account or API key.
Turns any URL into SEO metadata, contacts, tech stack, and AI-ready Markdown, in one call.
Fetch public webpages as clean text, Markdown, links, and metadata, with browser rendering.
Latest Blog Posts
- Who's Calling? MCP Hosts Are an Identity Blind Spot (And the Spec Knows It)By Om-Shree-0709 on .mcpAgent IdentityOAuth 2.1
- Your AI Chatbot Just Exposed Your CEO's Salary to an InternBy Om-Shree-0709 on .Agent IdentityMCP SecurityOAuth Delegation
- Why MCP Servers Need Execution Sandboxing (And Why Your Current Stack Isn't Enough)By Om-Shree-0709 on .Agentic AiPrompt InjectionWebAssembly
MCP directory API
We provide all the information about MCP servers via our MCP API.
curl -X GET 'https://glama.ai/api/mcp/v1/servers/dev55acc-ai/website-content-crawler-mcp'
If you have feedback or need assistance with the MCP directory API, please join our Discord server