Skip to main content
Glama
83,310 servers. Updated

Matching MCP tools:

Matching MCP Connectors:

"Web scraping and content extraction" matching MCP servers:

GET /v1/servers — MCP directory API reference
  • A
    license
    A
    quality
    C
    maintenance
    Enables AI agents to extract clean, structured web content (articles, tables, links, visual layouts) optimized for LLM token efficiency, with fast response times and optional JavaScript support.
    5
    64
    MIT
  • F
    license
    A
    quality
    D
    maintenance
    Enables retrieval and cleaning of official documentation content for popular AI/Python libraries (uv, langchain, openai, llama-index) through web scraping and LLM-powered content extraction. Uses Serper API for search and Groq API to clean HTML into readable text with source attribution.
    1
    4
    -
  • F
    license
    Not graded
    quality
    B
    maintenance
    Enables running the lead scoring and outbound pipeline by talking to Claude, including scraping healthcare job postings, scoring against buyer ICP, gating on company size, and managing sequences.
    -
  • A
    license
    A
    quality
    A
    maintenance
    Extract content from URLs, documents, videos, and audio files using intelligent auto-engine selection. Supports web pages, PDFs, Word docs, YouTube transcripts, and more with structured JSON responses.
    2
    171
    MIT
  • F
    license
    Not graded
    quality
    D
    maintenance
    Enables LLMs to extract content from websites using automated static and dynamic scraping engines with built-in anti-bot protections. It provides tools for web data retrieval and stores results in MongoDB with support for JSON and CSV exports.
    -
  • A
    license
    A
    quality
    A
    maintenance
    Provides LLMs with real-time web search and content extraction capabilities, including text/news search, full-text URL reading, and targeted technical documentation search.
    3
    24
    MIT
  • F
    license
    A
    quality
    C
    maintenance
    Enables live-web research by providing stateless browsing, link extraction, and page search capabilities via Playwright, allowing agents to fetch and analyze web content without session management.
    3
    -
  • F
    license
    A
    quality
    C
    maintenance
    Enables a locally-run LLM to search the web, fetch pages as markdown, make arbitrary HTTP requests, and optionally render pages with headless Chromium.
    3
    -
  • A
    license
    A
    quality
    B
    maintenance
    Provides an LLM client with self-hosted web scraping capabilities including scraping, crawling, mapping, searching, structured data extraction, browser interaction, and deep research, without reliance on paid proxy or CAPTCHA services.
    4
    GPL 3.0
  • A
    license
    A
    quality
    D
    maintenance
    Enables Claude and other LLMs to read and parse web content, with smart fallback strategies to bypass access restrictions (e.g., paywalls, Cloudflare) and output Markdown.
    5
    22
    44
    MIT
  • A
    license
    A
    quality
    C
    maintenance
    Provides DuckDuckGo web search and site fetching tools for Claude Desktop, enabling up to 10 search results and text content extraction from URLs.
    2
    MIT
  • A
    license
    A
    quality
    D
    maintenance
    Provides Baidu hot search and Bilibili trending video data for Claude and Cherry Studio, supporting retrieval, search, and cache management.
    6
    22
    1
    MIT