Skip to main content
Glama

web-crawler-mcp

A minimal crawl4ai crawler exposed as an MCP server, deployable to Cloud Run.

What it does

Exposes a single MCP tool, crawl(url, max_length), which fetches a page with crawl4ai's headless-browser crawler and returns its content as markdown.

Related MCP server: crawl-mcp-server

Local development

pip install -r requirements.txt
playwright install --with-deps chromium
crawl4ai-setup
python src/server.py

The server listens on PORT (default 8080) using the MCP Streamable HTTP transport, reachable at http://localhost:8080/mcp.

Deployment

.github/workflows/deploy.yml builds the Docker image with Cloud Build and deploys it to Cloud Run on every push to main that touches src/, Dockerfile, or requirements.txt.

Required GitHub Actions secrets:

  • GCP_SA_KEY — service account key JSON used to authenticate to Google Cloud.

  • GCP_PROJECT_ID — the target GCP project ID.

The Cloud Run service (web-crawler-mcp, region us-central1) is deployed with --allow-unauthenticated, matching the reference deployment pattern. Restrict access with IAM invoker bindings or a load balancer in front if the endpoint needs to be private.

Connecting an MCP client

Point an MCP client (Streamable HTTP transport) at:

https://<cloud-run-service-url>/mcp

Maintenance

ActivityMaintained
ResponsivenessSyncing

Related MCP Connectors

Related MCP Servers

  • F
    license
    Not graded
    quality
    Not graded
    maintenance
    An MCP server for web content extraction that converts HTML pages into clean, LLM-optimized Markdown using Mozilla's Readability. It supports batch processing, intelligent multi-page crawling, and configurable caching while respecting robots.txt standards.
    15
    -
  • A
    license
    Not graded
    quality
    D
    maintenance
    A lightweight MCP server that exposes Crawl4AI web scraping and crawling capabilities as tools for AI agents, enabling single-page scraping and multi-page crawling with adaptive stopping.
    108
    MIT