Skip to main content
Glama
619,810 tools. Updated 2026-09-28 18:01

"Scraping Public Documents" matching MCP tools:

  • Run a configured Spec pipeline on workspace documents or public file URLs to extract structured rows in one call.
    MIT
  • Route requests through datacenter IPs for high-volume scraping of APIs and public data feeds without aggressive anti-bot protections. Supports geo-targeting and sticky IP sessions.
    MIT
  • Search the web to find relevant pages, returning URLs, titles, and snippets immediately for pre-scraping discovery.
    Apache 2.0

Matching MCP Servers

Matching MCP Connectors

  • Crawl a website to discover and scrape multiple pages, following links up to a set depth. Filter URLs with regex and use saved or custom scraping configs.
    MIT
  • Browse files stored in your workspace. Filter by type to find specific assets like fonts, images, or documents, and retrieve public URLs for each file.
    MIT
  • Add a product to your shop catalog using your own data, with no scraping needed. Specify details like title, price, and affiliate URL to create the entry directly.
    MIT
  • Extract clean content from a public URL, returning Markdown, text, metadata, links, JSON-LD schema, and sections for scraping, RAG ingestion, or content analysis.
    MIT
  • Insert new documents or update existing ones by ID in a Zvec collection. Modify documents without checking existence first.
    Apache 2.0
  • Search your published PasteAI documents by entering a keyword to find documents with matching titles.
    MIT
    Destructive
  • Retrieve a list of recently published documents from PasteAI. Quickly see shared documents and access them.
    MIT
    Destructive
  • Add documents to a Meilisearch index using the MCP server. Specify indexUid and documents array to manage data efficiently for search and retrieval.
    MIT
  • Retrieve documents from a specified Meilisearch index using indexUid, with optional offset and limit parameters for precise data extraction.
    MIT