Skip to main content
Glama
306,462 tools. Last updated 2026-07-25 10:29

"Crawling Websites to Extract Data" matching MCP tools:

  • Extract web page content and convert it to clean, readable markdown format for analysis, bypassing paywalls and obtaining structured text data from websites.
    Apache 2.0
  • Extract structured financial data from investor relations websites and online sources for investment research when APIs are unavailable.
    MIT
  • Extract web page content and convert it to clean markdown format for reading articles, documentation, or analyzing text from websites.
    Apache 2.0
  • Extract, summarize, or scrape web content from URLs to read articles, crawl sites, or extract structured data using multiple providers.
    MIT
  • Crawl websites with nested URL navigation to extract and convert content into clean markdown format for structured data analysis.
    MIT

Matching MCP Servers

  • -
    license
    -
    quality
    B
    maintenance
    Enables OCR text extraction from images via URLs or base64, with pay-per-call using x402 micropayments (USDC on Base L2).
    Last updated

Matching MCP Connectors

  • Web content extraction for AI agents. Pay per call with x402 (USDC on Base). No API key.

  • URL to clean article markdown/text + metadata and links. Deterministic. $0.001/call via x402.

  • Recursively crawl websites to extract content in markdown, text, or raw formats, with configurable depth and page limits for data collection.
    MIT
  • Analyze competitor websites to extract SEO metrics, content insights, and performance data for competitive research and strategy development.
    MIT
  • Extract content from websites by crawling multiple pages from a starting URL, with configurable depth and page limits for structured data collection.
  • Perform AI-powered web searches to extract structured data from search results for research, competitive analysis, and multi-source information gathering.
    MIT
  • Search the web for current information, news, articles, and websites to find up-to-date content, research topics, or answer questions about recent events.
    Apache 2.0
  • Retrieve a list of all websites in your workspace. View available sites to manage or select one for further actions.
    MIT
  • Retrieve a project's saved configuration including base URL, max pages, screenshot directory, and authentication method. Use it to verify setup before crawling or testing.
    AGPL 3.0
  • Extract data from Excel sheets and convert to structured JSON with headers and typed rows. Specify sheet index and header presence for precise data extraction.
    MIT
  • Retrieve a paginated list of websites owned by your account on Kleap. Use offset and limit to control results.
    MIT