Skip to main content
Glama
465,476 tools. Updated 2026-08-19 03:55

"MCP server for web scraping and content extraction" matching MCP tools:

  • Detect changes on web pages over time. Extract content, compute hash, and compare with previous check to identify price updates, stock changes, or content modifications.
    MIT
  • Get instructions for obtaining an API key and using a web-access gateway that connects to 24 search, scraping, browser, and voice providers at provider-metered prices.
    MIT
  • Extract and process web content from URLs for data collection, content analysis, and research tasks, supporting multiple formats and extraction depths.
    MIT
  • Extract specific structured data from web pages using LLM-powered extraction. Use a prompt or schema to get details like prices, names, and descriptions.
    MIT

Matching MCP Servers

  • A
    license
    A
    quality
    C
    maintenance
    Enables AI agents to extract clean, structured web content (articles, tables, links, visual layouts) optimized for LLM token efficiency, with fast response times and optional JavaScript support.
    5
    81
    MIT
  • F
    license
    A
    quality
    D
    maintenance
    Enables retrieval and cleaning of official documentation content for popular AI/Python libraries (uv, langchain, openai, llama-index) through web scraping and LLM-powered content extraction. Uses Serper API for search and Groq API to clean HTML into readable text with source attribution.
    1
    2

Matching MCP Connectors

  • Generic URL crawl + HTML extraction — fallback for sites without dedicated MCPs.

  • 40+ web scraping tools from Firecrawl, Bright Data, Jina, Olostep, ScrapeGraph, Notte, and Riveter. Scrape, crawl, screenshot, and extract from any website. Starts at $0.01/call. Get your API key at app.xpay.sh or xpay.tools

  • Retrieve account subscription details, usage statistics, and quota limits for monitoring billing and concurrency management within the Scrapfly MCP server.
    ISC
  • Fetch any public web page and extract its readable content as clean markdown. Respects robots.txt, handles redirects, ideal for research and RAG.
    MIT
  • Perform comprehensive web searches with full page content extraction for detailed research and analysis across multiple sources.
    MIT
  • Extract structured data from web content using AI. Specify a URL and extraction instructions to retrieve organized information from articles, documents, or text-based web pages.
    MIT
  • Extract clean readable content from up to 20 URLs: text, title, author, date. Eliminates scraping. Feed web pages into AI agent context windows.
    MIT
  • Extract raw HTML from any URL for data extraction, content analysis, or price monitoring. Automatically handles anti-bot protection.
    MIT
  • Retrieve and extract clean text from web pages and PDFs, including JavaScript-rendered content, with automatic fallback to alternative extraction methods.
    MIT
  • Extract full text content from any web page URL with customizable output formats and options for metadata, tables, comments, and images.
    MIT
  • Extract the current web page's readable content as markdown for summarization or data extraction.
    MIT
  • Scrape content from a single web page with advanced options, supporting markdown, JSON schema-based extraction, and brand identity analysis.
    MIT
  • Extract content from URLs or local files, including web pages, PDFs, documents, and YouTube transcripts, without an API key for most sources. For audio/video transcription, an API key is required.
    MIT