Web Scraping
Specialized tools for extracting data from websites in a structured format. Enables automated collection of web content through headless browsers and parsing techniques.
MCP ServersBrowse all →
AlicenseAqualityCmaintenance42 local SEO tools for AI assistants. SERP tracking, Google Business Profile data, review monitoring, keyword research, AI visibility scoring, geogrid rank scans, citation audits, and competitive analysis.84218 npm3MIT
AllNewsAPI MCPofficial
AlicenseAqualityAmaintenanceGet access to real-time and historical news data including top headlines from global sources via AllNewsAPI. Supports multiple filter options including keyword search, category, language and more431,109 npm1MIT
HasData MCP Serverofficial
AlicenseAqualityAmaintenanceDirect access to 40+ scraping and search tools. Extract structured data from Google (Search, Maps, Trends), Amazon, Airbnb, Social Media, and any web page directly into your AI agent.8636MIT- AlicenseAqualityAmaintenanceEnables fetching YouTube transcripts, searching videos, and listing channel/playlist videos via the TranscriptFetch API.125268 npmMIT
- AlicenseAqualityCmaintenanceConverts raw HTML into structured, AI-readable page maps with 97% token reduction, enabling agents to read, click, type, and navigate any web page.91336AGPL 3.0

Yellow Pages MCP Serverofficial
AlicenseAqualityBmaintenanceEnables MCP clients to search Yellow Pages local businesses by keyword and location and retrieve full listing details such as phone, hours, services, brands, and payment methods.22109 npm15MIT
HatFetchofficial
AlicenseAqualityAmaintenanceEnables LLM agents to read any website by scraping and crawling into clean Markdown, automatically bypassing bot detection with residential proxies.239 npmMIT
Playwright MCP Serverofficial
AlicenseBqualityAmaintenanceA Model Context Protocol server that enables LLMs to interact with web pages through structured accessibility snapshots without requiring vision models or screenshots.10264,633,135 npm37,139Apache 2.0
SeleniumBase MCPofficial
AlicenseAqualityAmaintenanceBrowser automation that bypasses anti-bot systems and handles web-scraping with CDP Mode.5772413,019MIT
seleniumbase-mcpofficial
AlicenseAqualityAmaintenanceEnables MCP clients like Claude Desktop and Claude Code to drive a real browser with SeleniumBase automation. Offers three variants: WebDriver-based general automation, Pure CDP Mode for bot-detection bypass, and a broad SB() API with MFA and downloads.126248MIT- AlicenseAqualityDmaintenanceChecks username availability and retrieves account information across major social media platforms including YouTube, TikTok, Threads, X (Twitter), and Instagram, perfect for brand name qualification and competitive research.266 npmMIT
- AlicenseAqualityCmaintenanceMCP server for browser automation that lets LLMs interact with web pages through structured accessibility snapshots, bypassing the need for screenshots.8234,633,135 npmApache 2.0
- AlicenseBqualityDmaintenanceA Model Context Protocol server that provides browser automation capabilities using BrowserCat's cloud browser service. This server enables LLMs to interact with web pages, take screenshots, and execute JavaScript in a real browser environment without needing to install browsers locally.74 npm6MIT

agentminds-mcpofficial
AlicenseAqualityCmaintenanceMCP server for AgentMinds collective intelligence platform, enabling AI agents to scan websites for security/SEO/performance issues, pull personalized recommendations, and share findings across the network.74 npmMIT
anyapi-mcpofficial
AlicenseAqualityBmaintenanceLocal stdio MCP server for AnyAPI - hundreds of scraping and data APIs behind one key, priced per request in USD.10849 npm1Apache 2.0
TerraDeed MCP Serverofficial
AlicenseAqualityBmaintenanceEnables web scraping and structured data extraction via x402-powered USDC payments on Base mainnet, with tools for scraping URLs into markdown, extracting schema-defined JSON, and checking wallet balances.317 npmMIT
CubicJ MCP Searchofficial
AlicenseAqualityBmaintenanceEnables LLM clients to run Brave LLM Context web searches and fetch readable markdown extracts from URLs, including specialized site extraction, PDF text, raw mode, and pagination.2MIT- AlicenseAqualityBmaintenanceA local MCP server that leverages your real Chrome session to read and interact with web pages, including JavaScript-rendered and login-required content. Supports page actions like clicking, scrolling, typing, and platform-specific tools for Zhihu, Xiaohongshu, X, Reddit, and search engines.4246MIT
- AlicenseAqualityAmaintenanceEnables MCP clients to search stays by destination and dates with rich filters and read full property details as structured JSON, without needing a Booking.com account or self-hosting.42515 npm18MIT
- AlicenseAqualityAmaintenanceEnables MCP clients to access public TikTok profiles, videos, comments, and keyword search as structured JSON over streamable HTTP, without a TikTok developer account or OAuth.64490 npm15MIT
- AlicenseAqualityCmaintenanceEnables capturing HTTP/HTTPS webpages as images or PDFs using Screenshot Scout, with support for full-page, element, and device-specific captures, page interactions, and caching.139 npmMIT

apexapi-mcpofficial
AlicenseAqualityCmaintenanceEnables calling 120+ AI models and reading live web pages (scrape, crawl, structured extract) from any MCP client using one API key.940 npmMIT
Oxylabs MCP Serverofficial
AlicenseBqualityCmaintenanceA scraper tool that leverages the Oxylabs Web Scraper API to fetch and process web content with flexible options for parsing and rendering pages, enabling efficient content extraction from complex websites.103,044 PyPI105MIT- AlicenseAqualityAmaintenanceSelf-hosted anti-detect browser for AI agents. Rust MCP runtime + own Firefox engine fork (C++-level stealth). page_a11y: semantic snapshots pierce shadow DOM, detect login_state, danger zones, and prompt injection. Act by ref — no CSS selectors. Evidence recording, Android personas, 20 tools, 6 platforms.419MIT

Yutori MCPofficial
AlicenseAqualityBmaintenanceMCP server enabling web monitoring, deep research, and browser automation through Yutori's web agentic technology.132,375 PyPI24Apache 2.0- AlicenseAqualityAmaintenanceProvides MCP tools for web scraping, including smart fetching, proxy rotation, and browser rendering to bypass anti-bot challenges and render JavaScript-heavy sites.4108 npm4MIT
- AlicenseAqualityBmaintenanceEnables querying DNS Shop product data via MCP: search, product details, reviews, specifications, variants, categories, brand catalogs, availability by city, and store listings.1218MIT

octen-mcpofficial
AlicenseAqualityAmaintenanceMCP server for Octen Extract that turns any URL into clean, LLM-ready markdown, with highlights and classification. Works with Claude, Cursor, VS Code, and other MCP clients.6108 npm16MIT
stripfeed-mcp-serverofficial
AlicenseAqualityBmaintenanceConverts any URL to clean, token-efficient Markdown for AI agents. Strips ads, navigation, and scripts. Supports CSS selectors, batch processing (10 URLs), token counting, and smart caching.33 npmMIT
floppydata-mcpofficial
AlicenseAqualityDmaintenanceLets agents fetch web data and manage rotating proxies through the Floppydata API v2 directly from MCP clients.75 npmMIT
MCP ConnectorsBrowse all →
An MCP server that gives any LLM or agent clean YouTube transcripts on demand: a single video, a whole channel, or a playlist, plus AI cleanup of auto-generated captions. API-key auth, credit-based, same backend as the public v1 API. Get a free API key with 25 free credits at youtubetranscriptdownload.com/account.
Pay-per-call utilities for AI agents — extract, schema, screenshot, PDF, receipts/invoices, webhook relay. Prepaid credits or x402 USDC on Base.
Hundreds of scraping and data APIs through one gateway — one key, USD pay-per-request, normalized schemas, automatic failover.
Micro-settled web content extraction and JSON schema validation utilities for AI agents and MCP clients, backed by the Bristlecone Logic API engine.
Convert PDF, DOCX, HTML, and URLs to clean, LLM-ready markdown with tables preserved
Website QA for your coding agent: audit SEO, performance, security, accessibility over MCP.
Search YouTube and read video, channel and transcript data as JSON. No Google Cloud project.
Paid APIs and tokenized shares for agents. Prepaid funding, including authorized Instinct checkout.
If it's here, it's on their board - live jobs from employer career pages, apply on theyond.com.
Docs: https://docs.keenable.ai/mcp-server Keenable is a free, remote MCP server that gives agents access to the web index. Search the web with ranked results and date/site filters, then fetch any indexed page as clean markdown. Works out of the box with no account or API key.
Indeed job listings by keyword and location, and full postings, as structured JSON.
DuckDuckGo search as JSON. Ranked results, ads in their own array, 37 regions.
Amazon keyword search, product details, seller profiles and seller catalogues, as structured JSON.
Yelp business search, place details and the full review feed, as structured JSON.
Yellow Pages local business search and full business listings, as structured JSON.
The public product catalogue and collections of any Shopify storefront, as structured JSON.
Public Instagram profiles and post feeds by handle, with hashtags and mentions parsed.
Google SERP as JSON: organic, AI Overview, People Also Ask, AI Mode, news, shopping. No Cloud setup.
Google Images results with source pages, thumbnails and full-resolution URLs, as JSON.
Booking.com stays by destination and dates, and full property details, as structured JSON.