Skip to main content
Glama
306,516 tools. Last updated 2026-07-25 11:18

"Tools for simulating browser behavior to scrape web pages" matching MCP tools:

  • Discover all indexed URLs on a website to identify pages for scraping or locate specific content when scrape results are incomplete.
    MIT
  • Discover all indexed URLs on a website. Use the search parameter to find specific pages when scrape returns empty or incomplete results.
    MIT
  • Extract structured data from web pages including markdown content, links, tables, images, and metadata. Supports JavaScript-rendered pages with optional CSS selector waiting for dynamic content.
    MIT
  • Scrape a single web page and extract content in markdown, text, or HTML. Strip navigation and ads, include links, and choose between HTTP or browser rendering.
    MIT
  • Given a URL, extracts the page title, readable markdown, and plain text. Helps AI agents read web content, scrape data, and feed RAG pipelines.
    MIT

Matching MCP Servers

  • -
    license
    -
    quality
    -
    maintenance
    Enables scraping and fetching websites with protection handling like Cloudflare and captchas, via an MCP interface.
    Last updated
    1
  • A
    license
    B
    quality
    D
    maintenance
    A comprehensive web scraping server that transforms web content into clean, agent-ready Markdown with automatic citations and efficient caching. It features a robust suite of tools for metadata extraction, sentiment analysis, SEO auditing, and security scanning while strictly adhering to robots.txt policies.
    Last updated
    48
    28
    29
    MIT

Matching MCP Connectors

  • x402-gated web extraction gateway. Tools: extract, extract_batch.

  • tldr-pages community simplified man pages (cached 24h)

  • Screenshot web pages for visual reconnaissance and technology identification using headless Chrome. Supports concurrent captures and custom resolution.
    MIT
  • Retrieve stored sessions with cookies for authenticated web scraping. Use the returned session IDs to scrape login-protected pages like Amazon orders or LinkedIn profiles.
    MIT
  • Open a new browser session to audit web pages, including those behind logins or multi-step journeys. Returns a session ID for navigation and interaction tools.
    MIT
  • Trigger hover events on web elements to reveal dropdown menus, tooltips, and hidden UI content by simulating realistic browser mouse behavior.
    MIT
  • Extract content from web pages using JavaScript rendering, antibot bypass, and captcha solving to access protected sites and SPAs.
    MIT
  • Search the web, scrape top results, and produce an aggregated cited report.
    MIT
  • Extract and scrape webpage content using auto, simple, Scrapy, or Selenium methods. Define extraction rules or wait for specific elements to retrieve targeted data.
    MIT
  • Retrieve content from JavaScript-heavy web pages using browser automation, with automatic cookie management and CSS selector extraction for targeted results.
    MIT
  • Automate web application testing with Playwright browser automation. Load and interact with pages in real browsers to detect runtime errors, hydration issues, and client-side problems.
    MIT
  • Perform real-time web searches with configurable parameters to retrieve and scrape content from relevant websites for AI assistants.
    MIT