Skip to main content
Glama
dev55acc-ai

website-content-crawler

by dev55acc-ai

Website Content Crawler

Crawls a list of URLs and returns page content as structured JSON. Runs as an Apify Actor and as an MCP server so agents can call it directly.

Live demo: https://website-content-crawler.vercel.app

What it costs to run

Pay per Event on Apify:

Event

Price

Actor start

$0.00005

Page record returned

$0.001

Platform compute is deducted before payout; see PRICING.md for unit economics.

Related MCP server: cleanfetch

Output

Every run returns this exact shape. The output schema enforces it.

{
  "url": "https://example.com",
  "status": 200,
  "title": "Example Domain",
  "description": null,
  "h1": "Example Domain",
  "markdown": "# Example Domain\n\nThis domain is for use in documentation examples without needing permission. Avoid use in operations.\n\nLearn more\n\nLearn more (https://iana.org/domains/example)",
  "links": [
    "https://iana.org/domains/example"
  ],
  "wordCount": 23,
  "crawledAt": "2026-08-23T09:55:20.140Z",
  "error": null
}

Layout

  • src/ — actor source (Crawlee Cheerio crawler)

  • .actor/ — Apify actor spec: actor.json, INPUT_SCHEMA.json, output_schema.json, dataset_schema.json, Dockerfile

  • mcp-server/ — MCP wrapper exposing the actor as the crawl_website tool

  • site/ + tools/ — builds the live demo page from a real run record

Run locally

npm install
npm test

MCP server:

cd mcp-server && npm install && node smoke.mjs

Publishing to the Apify Store

When the account is active: apify push, price as Pay per Event, listing uses this README.

A
license - permissive license
Not graded
quality - not tested
C
maintenance

Maintenance

Maintainers
Response time
Release cycle
Releases (12mo)
Commit activity

Resources

Unclaimed servers have limited discoverability.

Looking for Admin?

If you are the server author, to access and configure the admin panel.

Related MCP Servers

  • A
    license
    B
    quality
    D
    maintenance
    Enables LLMs to retrieve and process web content by fetching URLs and converting HTML to markdown format. Supports chunked reading of large pages and can access both public websites and local networks.
    1
    MIT
  • A
    license
    A
    quality
    B
    maintenance
    Enables AI agents to read web pages reliably, returning clean markdown content, hyperlinks, and metadata without navigation or ad noise.
    3
    14
    MIT
  • A
    license
    Not graded
    quality
    C
    maintenance
    Enables AI agents to fetch and extract clean, readable content from web pages, and search within pages for specific queries, without needing a full browser.
    1
    MIT

View all related MCP servers

Related MCP Connectors

  • Read a URL as clean markdown, screenshot a website, url to PDF. Web access for agents, no signup.

  • Web scraping for AI agents. Converts URLs to clean, LLM-ready Markdown with anti-bot bypass.

  • Read any web page as clean Markdown for AI agents: fetch, search, metadata, links. SSRF-safe.

View all MCP Connectors

Latest Blog Posts

MCP directory API

We provide all the information about MCP servers via our MCP API.

curl -X GET 'https://glama.ai/api/mcp/v1/servers/dev55acc-ai/website-content-crawler'

If you have feedback or need assistance with the MCP directory API, please join our Discord server