Skip to main content
Glama
83,309 servers. Updated

Matching MCP tools:

Matching MCP Connectors:

"Information about extraction techniques or processes" matching MCP servers:

GET /v1/servers — MCP directory API reference
  • A
    license
    B
    quality
    D
    maintenance
    Provides web fetching and content extraction tools (fetch_html, extract_readable, query_selector) and a tool to retrieve Honor of Kings hero skin cover images by hero name.
    4
    MIT
  • A
    license
    A
    quality
    A
    maintenance
    Local MCP server for web search and page extraction, providing clean markdown from URLs, search results, site mapping, and research endpoints without API keys or accounts.
    5
    Apache 2.0
  • A
    license
    A
    quality
    B
    maintenance
    Scrape any webpage and return clean markdown, HTML, or structured JSON. Bypasses anti-bot protection, renders JavaScript (React/Vue/Angular), supports premium residential proxies and CSS extraction. Works with any MCP client — no local install required.
    1
    286
    19
    MIT
  • A
    license
    A
    quality
    D
    maintenance
    Integrates Firecrawl for web scraping, crawling, search, and content extraction capabilities. Supports single/batch scraping, URL discovery, structured data extraction, deep research, and AI-powered web analysis with automatic retries and rate limiting.
    8
    28,853
    MIT
  • A
    license
    A
    quality
    D
    maintenance
    Integrates Firecrawl web scraping capabilities, enabling web content extraction, crawling, site mapping, search, and structured data extraction with automatic rate limiting and retry handling.
    6
    28,853
    2
    MIT
  • A
    license
    A
    quality
    D
    maintenance
    A headless web-scraping MCP server built on Scrapy, providing tools for polite fetching, CSS/XPath extraction, link/table extraction, sitemap and robots.txt reading, and bounded asynchronous crawls.
    10
    MIT
  • A
    license
    A
    quality
    B
    maintenance
    Fast, token-efficient web content extraction tool that converts websites to clean Markdown for AI agents, featuring smart caching, content extraction with Mozilla Readability, and polite crawling capabilities.
    1
    534
    161
    MIT
  • A
    license
    A
    quality
    A
    maintenance
    A privacy-friendly MCP server that enables web searches and URL content extraction using DuckDuckGo, allowing AI assistants to access real-time web information without API keys.
    1
    665
    42
    Apache 2.0
  • A
    license
    A
    quality
    A
    maintenance
    Enables AI agents to search and fetch high-quality information from multiple sources, including general web APIs, Wikipedia, arXiv, Hacker News, Stack Exchange, and Crossref, with optional pro-mode deep research and clean markdown page extraction.
    10
    MIT
  • -
    license
    A
    quality
    Not graded
    maintenance
    Enables web search through Google and Wikipedia plus content extraction from any webpage via the Seekr API. Provides real-time search results with advanced filtering options and clean text extraction capabilities.
    2
    -
  • A
    license
    A
    quality
    D
    maintenance
    Enables web content scanning and analysis by fetching, analyzing, and extracting information from web pages using tools like page fetching, link extraction, site crawling, and more.
    6
    13
    MIT
  • A
    license
    A
    quality
    B
    maintenance
    An agent-agnostic web extraction and fetch layer that turns URLs into verified, typed data with confidence scores via MCP, REST, or SDK, orchestrating scraping engines behind a resilience ladder and supporting structured extraction against any schema.
    5
    2
    Apache 2.0
  • A
    license
    A
    quality
    B
    maintenance
    Provides an LLM client with self-hosted web scraping capabilities including scraping, crawling, mapping, searching, structured data extraction, browser interaction, and deep research, without reliance on paid proxy or CAPTCHA services.
    4
    GPL 3.0
  • A
    license
    A
    quality
    C
    maintenance
    Provides access to Jina AI's suite of web tools including URL reading, web/image/academic search, content extraction, embeddings, and reranking capabilities. Enables AI assistants to search the web, extract content from URLs, and process information with semantic understanding.
    19
    Apache 2.0
  • A
    license
    A
    quality
    F
    maintenance
    A powerful web scraping MCP server built on Scrapy and FastMCP that supports multiple scraping methods (HTTP, Scrapy, browser automation), anti-detection techniques, form handling, and concurrent crawling. Designed for commercial environments with enterprise-grade features like intelligent retry mechanisms, performance monitoring, and configurable data extraction.
    10
    3
    MIT
  • F
    license
    A
    quality
    D
    maintenance
    Provides AI-powered web search, full-page content extraction, and search-enhanced Q\&A capabilities via the Metaso AI search engine. It enables large language models to access diverse information across web, academic, and multimedia sources with structured Markdown or JSON output.
    3
    1
    -
  • A
    license
    A
    quality
    A
    maintenance
    AgisHub provides a diverse set of pay-per-call tools for AI agents—covering timezones, web extraction, AI, memory, webhooks, finance, and utilities—accessible via a free MCP tier or x402 HTTP micropayments in USDC on Base, with no signup or API keys required.
    25
    89
    1
    MIT
  • A
    license
    A
    quality
    D
    maintenance
    Enables comprehensive web and news searches via the Google Custom Search API with integrated content extraction using the Mozilla Readability algorithm. It allows users to perform quick snippet lookups or deep searches that fetch and format full article content into clean markdown.
    3
    12
    2
    MIT
  • A
    license
    A
    quality
    C
    maintenance
    Enables AI agents and MCP clients to automate web browsers via Selenium WebDriver, supporting Chrome, Firefox, and Edge in headless or visible mode with tools for navigation, interaction, content extraction, screenshots, and scripting.
    21
    13
    MIT