Skip to main content
Glama
dev55acc-ai

Website Content Crawler MCP Server

by dev55acc-ai

Website Content Crawler MCP Server

MCP 包装器,通过 apify-client 真实运行 apify/website-content-crawler actor,并将其页面以 JSON 形式返回。每次调用要么返回抓取的内容,要么返回结构化错误——绝不会报告虚假成功。

费用

运行费用按 apify/website-content-crawler 页面列出的费率计入你的 Apify 账户——此服务器不额外加价。没有 token 就不会产生费用:任何运行开始前,调用都会返回 missing_token

实时输出演示(相同的抓取逻辑,渲染效果):https://website-content-crawler.vercel.app

Related MCP server: Crawl4AI MCP Server

设置

npm install
export APIFY_TOKEN=apify_api_...   # https://console.apify.com/settings/integrations
npm start                          # stdio MCP server

工具:crawl_website

输入参数:

字段

类型

默认值

说明

url

string

必填

要抓取的 http/https URL

maxPages

number

10

上限 50

crawlerType

string

cheerio

playwright:chrome(用于 JS 渲染页面)

输出信封(每次调用的结构相同):

{
  "status": "ok",
  "run": { "id": "<apify run id>", "status": "SUCCEEDED" },
  "page_count": 3,
  "total_in_dataset": 3,
  "pages": [{ "url": "...", "title": "...", "text": "...(≤5000 chars)" }]
}

错误代码:invalid_urlmissing_tokenapify_auth_failedactor_run_failedrun_not_succeededdataset_fetch_failed

冒烟测试

printf '%s\n' \
 '{"jsonrpc":"2.0","id":1,"method":"initialize","params":{"protocolVersion":"2024-11-05","capabilities":{},"clientInfo":{"name":"t","version":"0"}}}' \
 '{"jsonrpc":"2.0","method":"notifications/initialized"}' \
 '{"jsonrpc":"2.0","id":2,"method":"tools/list"}' \
 '{"jsonrpc":"2.0","id":3,"method":"tools/call","params":{"name":"crawl_website","arguments":{"url":"https://example.com"}}}' \
 | node index.js

在未设置 APIFY_TOKEN 的情况下,请求 id 3 必须返回 {"status":"error","error":{"code":"missing_token",...}}——证明 handler 能到达真实 Apify 边界,而不是凭空编造结果。

Install Server
F
license - not found
A
quality
C
maintenance

Maintenance

Maintainers
Response time
Release cycle
Releases (12mo)
Commit activity

Resources

Unclaimed servers have limited discoverability.

Looking for Admin?

If you are the server author, to access and configure the admin panel.

Related MCP Servers

  • A
    license
    Not graded
    quality
    D
    maintenance
    Enables web scraping and document processing with JavaScript execution, anti-detection measures, batch processing, and structured data extraction. Supports multiple formats including markdown, HTML, screenshots, and handles PDFs with OCR capabilities.
    3
    MIT
  • F
    license
    Not graded
    quality
    D
    maintenance
    Enables advanced web crawling and content extraction with JavaScript support, AI-powered analysis, PDF/Office document processing, YouTube transcript extraction, Google search integration, and multi-format data export capabilities.
    2
  • F
    license
    Not graded
    quality
    D
    maintenance
    Provides web crawling and browser automation capabilities with support for multiple content formats (HTML, JSON, PDF, screenshots, Markdown), page content extraction, console message monitoring, and network request tracking.
  • A
    license
    A
    quality
    A
    maintenance
    Enables web scraping, structured data extraction, and screenshot capture with automatic anti-bot bypass, supporting JavaScript rendering, proxy rotation, and tiered pricing.
    25
    187
    1
    MIT

View all related MCP servers

Related MCP Connectors

  • Scrape, crawl, map and extract the web. Pay per call in USDC, no account or API key.

  • Turns any URL into SEO metadata, contacts, tech stack, and AI-ready Markdown, in one call.

  • Fetch public webpages as clean text, Markdown, links, and metadata, with browser rendering.

View all MCP Connectors

Latest Blog Posts

MCP directory API

We provide all the information about MCP servers via our MCP API.

curl -X GET 'https://glama.ai/api/mcp/v1/servers/dev55acc-ai/website-content-crawler-mcp'

If you have feedback or need assistance with the MCP directory API, please join our Discord server