Skip to main content
Glama
dev55acc-ai

website-content-crawler

by dev55acc-ai

Website Content Crawler

Rastrea una lista de URLs y devuelve el contenido de las páginas como JSON estructurado. Se ejecuta como un Apify Actor y como un servidor MCP para que los agentes puedan llamarlo directamente.

Demo en vivo: https://website-content-crawler.vercel.app

Lo que cuesta ejecutarlo

Pago por evento en Apify:

Evento

Precio

Inicio del actor

$0.00005

Registro de página devuelto

$0.001

La computación de la plataforma se descuenta antes del pago; consulta PRICING.md para conocer la economía unitaria.

Related MCP server: cleanfetch

Salida

Cada ejecución devuelve exactamente esta forma. El esquema de salida la exige.

{
  "url": "https://example.com",
  "status": 200,
  "title": "Example Domain",
  "description": null,
  "h1": "Example Domain",
  "markdown": "# Example Domain\n\nThis domain is for use in documentation examples without needing permission. Avoid use in operations.\n\nLearn more\n\nLearn more (https://iana.org/domains/example)",
  "links": [
    "https://iana.org/domains/example"
  ],
  "wordCount": 23,
  "crawledAt": "2026-08-23T09:55:20.140Z",
  "error": null
}

Estructura

  • src/ — código fuente del actor (rastreador Crawlee Cheerio)

  • .actor/ — especificación del actor de Apify: actor.json, INPUT_SCHEMA.json, output_schema.json, dataset_schema.json, Dockerfile

  • mcp-server/ — wrapper MCP que expone el actor como la herramienta crawl_website

  • site/ + tools/ — construyen la página de demo en vivo a partir de un registro de ejecución real

Ejecutar en local

npm install
npm test

Servidor MCP:

cd mcp-server && npm install && node smoke.mjs

Publicar en Apify Store

Cuando la cuenta esté activa: apify push, precio como Pago por evento, y el listado usa este README.

A
license - permissive license
Not graded
quality - not tested
C
maintenance

Maintenance

Maintainers
Response time
Release cycle
Releases (12mo)
Commit activity

Resources

Unclaimed servers have limited discoverability.

Looking for Admin?

If you are the server author, to access and configure the admin panel.

Related MCP Servers

  • A
    license
    B
    quality
    D
    maintenance
    Enables LLMs to retrieve and process web content by fetching URLs and converting HTML to markdown format. Supports chunked reading of large pages and can access both public websites and local networks.
    1
    MIT
  • A
    license
    A
    quality
    B
    maintenance
    Enables AI agents to read web pages reliably, returning clean markdown content, hyperlinks, and metadata without navigation or ad noise.
    3
    14
    MIT
  • A
    license
    Not graded
    quality
    C
    maintenance
    Enables AI agents to fetch and extract clean, readable content from web pages, and search within pages for specific queries, without needing a full browser.
    1
    MIT

View all related MCP servers

Related MCP Connectors

  • Read a URL as clean markdown, screenshot a website, url to PDF. Web access for agents, no signup.

  • Web scraping for AI agents. Converts URLs to clean, LLM-ready Markdown with anti-bot bypass.

  • Read any web page as clean Markdown for AI agents: fetch, search, metadata, links. SSRF-safe.

View all MCP Connectors

Latest Blog Posts

MCP directory API

We provide all the information about MCP servers via our MCP API.

curl -X GET 'https://glama.ai/api/mcp/v1/servers/dev55acc-ai/website-content-crawler'

If you have feedback or need assistance with the MCP directory API, please join our Discord server