Why this server?
This server directly addresses both parts of the request by integrating multiple search engines, including SearXNG, and providing scraping, crawling, and content extraction capabilities to process the resulting web links.
AlicenseAqualityAmaintenanceA Model Context Protocol server that enables web search, scraping, crawling, and content extraction through multiple engines including SearXNG, Firecrawl, and Tavily.488140MITWhy this server?
This solution explicitly links SearXNG search results with a tool (Puppeteer-scraper) capable of navigating and extracting live content from the identified web links, achieving the required workflow.
AlicenseAqualityDmaintenanceAn MCP server implementation that integrates the SearXNG API for powerful web search capabilities and uses @missionsquad/puppeteer-scraper to read and process live web content.2171MITWhy this server?
This server combines SearXNG search functionality with explicit capabilities for 'website content scraping,' making it suitable for finding links via SearXNG and then extracting the page content.
AlicenseNot gradedqualityDmaintenanceA Model Context Protocol server that enables web search with category support, website content scraping with citation metadata, and timezone-aware date/time tools.54MITWhy this server?
This general-purpose tool is ideal for the second step: crawling and extracting data from web pages given the links identified by SearXNG, outputting structured data in formats like Markdown.
AlicenseNot gradedqualityCmaintenanceEnables web scraping and crawling capabilities for LLM clients, supporting single-page scraping, multi-page website crawling, and web search with multiple engines (Playwright, Cheerio, Puppeteer) and flexible output formats including markdown, HTML, text, and screenshots.116MITWhy this server?
This tool explicitly supports both web search (finding the links) and subsequent crawling and content extraction (scraping the content from those links).
FlicenseCqualityCmaintenanceBuilt as a Model Context Protocol (MCP) server that provides advanced web search, content extraction, web crawling, and scraping capabilities using the Firecrawl API.41-Why this server?
Specializes in high-quality scraping and data extraction from any website globally, making it a robust option for accessing and extracting content from the pages linked by SearXNG.
-licenseNot gradedqualityNot gradedmaintenanceEnables AI models to scrape and extract data from any website globally using Thordata's 195+ country proxy network. Bypasses anti-bot systems and renders JavaScript content, outputting structured data in Markdown, HTML, or Links format.-Why this server?
This server provides the first necessary step: access to the SearXNG metasearch engine to identify the relevant web links.
AlicenseAqualityAmaintenanceAn MCP server implementation that integrates the SearxNG API, providing web search capabilities.214,6401,204MITWhy this server?
This tool is essential for the second step, enabling the fetching of content (HTML, text, etc.) from any given URL, such as the links returned by the SearXNG servers.
AlicenseAqualityDmaintenanceA Model Context Protocol (MCP) server that enables Claude or other LLMs to fetch content from URLs, supporting HTML, JSON, text, and images with configurable request parameters.33MITWhy this server?
Useful after fetching the content, this server extracts and cleans the main webpage content, transforming raw HTML into clean, organized Markdown for easy analysis.
AlicenseAqualityFmaintenanceExtracts and transforms webpage content into clean, LLM-optimized Markdown. Returns article title, main content, excerpt, byline and site name. Uses Mozilla's Readability algorithm to remove ads, navigation, footers and non-essential elements while preserving the core content structure.11217MIT