Skip to main content
Glama
519,265 tools. Updated 2026-09-06 07:24

"Web scraping and content extraction" matching MCP tools:

  • Search Apify and Crawlee docs via full-text queries to find relevant pages on platform features, SDKs, CLI, REST API, and web scraping libraries.
    MIT
  • Extract and process web content from URLs for data collection, content analysis, and research tasks, supporting multiple formats and extraction depths.
    MIT
  • Remove an extraction schema by scheme ID to delete outdated data extraction configurations and keep your scraping workflows clean.
    MIT
    Destructive
  • Fetch any public web page and extract its readable content as clean markdown. Respects robots.txt, handles redirects, ideal for research and RAG.
    MIT
  • Perform comprehensive web searches with full page content extraction for detailed research and analysis across multiple sources.
    MIT

Matching MCP Servers

  • A
    license
    A
    quality
    C
    maintenance
    Enables AI agents to extract clean, structured web content (articles, tables, links, visual layouts) optimized for LLM token efficiency, with fast response times and optional JavaScript support.
    5
    64
    MIT
  • F
    license
    Not graded
    quality
    C
    maintenance
    Extracts structured JSON from receipts and invoices with Australian GST/ABN validation, per-line tax codes, confidence scores, and rationale.
    2
    -

Matching MCP Connectors

  • AU GST/ABN receipt extraction — assigns entertainment/ITC tax codes per line, not just OCR.

  • Generic URL crawl + HTML extraction — fallback for sites without dedicated MCPs.