InfinityScrape MCP
Scans Codeforces for public user profiles and competitive programming activity as part of multi-platform OSINT reconnaissance.
Scans Dev.to for public developer profiles and posts as part of multi-platform OSINT reconnaissance.
Provides real-time web searches via DuckDuckGo, including text search and automatic scraping of top results.
Scans GitHub for public user, team, and repository information as part of multi-platform OSINT reconnaissance.
Scans GitLab for public user, group, and project information as part of multi-platform OSINT reconnaissance.
Scans Google Scholar for scholarly profiles and publications as part of multi-platform OSINT reconnaissance.
Scans Kaggle for public user profiles and dataset/competition activity as part of multi-platform OSINT reconnaissance.
Scans LeetCode for public user profiles and coding activity as part of multi-platform OSINT reconnaissance.
Scans Medium for public author profiles and articles as part of multi-platform OSINT reconnaissance.
Resolves addresses to GPS coordinates, street/postcode details, and administrative boundaries using OpenStreetMap.
Scans Reddit for public user profiles and activity as part of multi-platform OSINT reconnaissance.
Scans ResearchGate for public researcher profiles and publications as part of multi-platform OSINT reconnaissance.
Scans Substack for public author profiles and newsletters as part of multi-platform OSINT reconnaissance.
Extracts YouTube video, Shorts, and live transcripts with timestamps directly over HTTP, without requiring video downloads or local GPU models.
Click on "Install Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@InfinityScrape MCPScrape the content from https://example.com and summarize it, bypassing anti-bot checks."
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
๐ InfinityScrape MCP: World-Class Web Scraping & Deep OSINT Intelligence Suite
InfinityScrape MCP is a standalone, production-grade Model Context Protocol (MCP) server engineered to provide AI models (LM Studio, Claude Desktop, Cursor, Open WebUI, Antigravity AI) with unlimited, high-speed, anti-bot resilient web scraping, dynamic SPA rendering, instant YouTube transcription, and precision OSINT / GEOINT location intelligence.
๐ Table of Contents
Related MCP server: FineData MCP Server
๐ Why InfinityScrape MCP?
Standard web scrapers often fail on modern websites due to Cloudflare challenges, heavy client-side JavaScript rendering, intrusive cookie consent modals, and rate limits. InfinityScrape solves these problems out-of-the-box:
Dual-Engine Architecture:
Fast TLS Engine (
primp+httpx): Mimics real Chrome/Safari browser TLS/JA3 fingerprints and HTTP/2 headers to bypass Cloudflare and Akamai challenges in<100ms.Dynamic Headless Browser (
Playwright Chromium): Renders complex SPAs (React, Vue, Next.js, Angular), performs infinite scrolling, clicks elements, and executes custom JavaScript.
Network-Level Ad & Tracker Elimination:
Intercepts and aborts network calls to 35+ ad networks and tracking scripts (
doubleclick,criteo,outbrain,google-analytics) before they download, cutting page load time by ~300% and memory usage by 70%.Automatically detects and decomposes OneTrust, Cookiebot, and sticky overlay popups.
Zero-GPU Instant YouTube Transcriber:
Extracts complete video/shorts/live transcripts with timestamps (
[MM:SS]) in<300msdirectly via HTTP streams without downloading video or requiring local GPU Whisper models.
Deep Recursive Documentation Crawler:
Asynchronous Breadth-First-Search (BFS) crawler with domain locking and path prefix filtering to aggregate entire documentation trees into unified Markdown.
State-of-the-Art Public OSINT & GEOINT Reconnaissance:
Multi-Signal Confidence Scoring (0% - 100%): Evaluates Name + City + Street + PIN + Org + Role correlation to rank discovered dossiers.
25+ Global Platform Scanners: Scans GitHub, GitLab, StackOverflow, Kaggle, HuggingFace, LeetCode, Codeforces, Dev.to, Medium, Substack, Google Scholar, ResearchGate, Reddit, etc.
OpenStreetMap GEOINT: Resolves global addresses down to street/postcode level with GPS coordinates and administrative boundaries.
SQLite Persistent Caching Layer:
In-memory and SQLite-backed local cache for instant
0msresponses on repeat lookups with configurable TTL.
โก Competitive Comparison
Feature / Capability | Standard MCP Scrapers | Cloud Scraping APIs | InfinityScrape MCP |
Cost & API Keys | Free (Basic) | Paid ($20 - $200/mo) | 100% Free / Zero API Keys |
Cloudflare / Akamai TLS Bypass | โ Fails / 403 | โ Yes | โ
Built-in ( |
Dynamic SPAs & Infinite Scroll | โ Limited | โ Yes | โ
Built-in ( |
Network-Level Ad & Popup Stripping | โ No | โ ๏ธ Partial | โ Built-in (35+ domains) |
Zero-GPU YouTube Transcripts | โ No | โ No | โ Built-in (<300ms) |
Online PDF Page-by-Page Parser | โ No | โ ๏ธ Extra Cost | โ
Built-in ( |
Deep Documentation Crawler | โ No | โ ๏ธ Extra Cost | โ Built-in (Async BFS) |
25+ Platform OSINT & Geocoding | โ No | โ No | โ Built-in (0-100% Confidence) |
Local SQLite 0ms Caching | โ No | โ No | โ Built-in (Auto TTL) |
๐๏ธ Architectural Overview
โโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโ
โ AI Client (LM Studio / Claude / Cursor) โ
โโโโโโโโโโโโโโโโโโโโโโโโโฌโโโโโโโโโโโโโโโโโโโโโโโโโ
โ JSON-RPC 2.0 (Stdio)
โผ
โโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโ
โ InfinityScrape MCP Server โ
โ (server.py) โ
โโโโโโโโโฌโโโโโโโโโโโโโโโโโฌโโโโโโโโโโโโโโโโฌโโโโโโโโ
โ โ โ
โโโโโโโโโโโโโโโโโโโดโโ โโโโโโโโโโดโโโโโโโโโ โโโดโโโโโโโโโโโโโโโโโ
โผ โผ โผ โผ โผ โผ
[Fast TLS Engine] [Playwright Engine] [OSINT / GEOINT] [Media & PDF Engines]
โข primp JA3/TLS โข Stealth Chromium โข 25+ Platform โข YouTube (<300ms)
โข HTTP/2 Stealth โข Network Ad Blocker Scanners โข Remote PDF Stream
โข <100ms Execution โข Infinite Scroll โข OpenStreetMap โข Table Markdownify
โข Auto-Dismiss CMPs โข Match Confidence
โ
โผ
โโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโ
โ SQLite Caching Layer (0ms TTL) โ
โโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโ๐ Quick Start & 1-Click Installation
Prerequisites
Python 3.10, 3.11, or 3.12+ installed.
Windows, macOS, or Linux.
1-Click Setup:
On Windows:
Double-click install.bat or run in PowerShell:
.\install.batOn Linux / macOS:
chmod +x install.sh
./install.shManual Setup (Any Platform):
# 1. Create virtual environment
python -m venv .venv
# 2. Activate virtual environment
# Windows: .venv\Scripts\activate | Linux/Mac: source .venv/bin/activate
# 3. Install requirements & Playwright browser
pip install -r requirements.txt
playwright install chromium๐ AI Client Integration
1. LM Studio (v0.3+)
Go to Settings โ Developer โ MCP Servers โ Edit Config and add:
{
"mcpServers": {
"infinity-scraper": {
"command": "C:/path/to/infinity-scraper/.venv/Scripts/python.exe",
"args": [
"-m",
"infinity_scraper.server"
],
"cwd": "C:/path/to/infinity-scraper",
"env": {
"PYTHONUNBUFFERED": "1"
}
}
}
}2. Claude Desktop
Edit %APPDATA%\Claude\claude_desktop_config.json (Windows) or ~/Library/Application Support/Claude/claude_desktop_config.json (macOS):
{
"mcpServers": {
"infinity-scraper": {
"command": "C:/path/to/infinity-scraper/.venv/Scripts/python.exe",
"args": [
"-m",
"infinity_scraper.server"
]
}
}
}3. Cursor IDE
In Cursor Settings โ Features โ MCP Servers โ Add New MCP Server:
Name:
infinity-scraperType:
commandCommand:
C:/path/to/infinity-scraper/.venv/Scripts/python.exe -m infinity_scraper.server
๐ ๏ธ Complete 23-Tool Reference Guide
1. Web Scraping & Content Crawling
Tool | Purpose | Key Parameters |
| Universal scraping with auto-upgrading TLS-to-Browser engine. |
|
| Headless browser for SPAs, infinite scrolls, and click actions. |
|
| Real-time internet search via DuckDuckGo. |
|
| Searches web and automatically scrapes top results into a cited report. |
|
| Recursive async BFS site crawler for full documentation trees. |
|
| Concurrently scrape multiple URLs in parallel. |
|
| Extract targeted fields using a CSS selector map to JSON. |
|
| Extract JSON-LD, OpenGraph metadata, and HTML tables. |
|
| Semantic RAG chunker & token optimizer for massive pages. |
|
2. Media, Social, Video & Document Parsers
Tool | Purpose | Key Parameters |
| Zero-GPU YouTube video transcript extraction with timestamps. |
|
| Streams and extracts remote online PDF documents page-by-page. |
|
| Extracts camera specs, timestamps, and GPS geotags from photos. |
|
| Ingests Reddit posts, scores, and nested comment dialogues. |
|
| Real-time RSS/Atom feed parser for blogs, Substack, and news. |
|
3. Deep Public OSINT & Entity Reconnaissance
Tool | Purpose | Key Parameters |
| Multi-domain open web profile scraper & confidence-ranked dossier builder. |
|
| Global OpenStreetMap forward geocoding & administrative breakdown. |
|
| Hierarchical location drill-down search matrix with negative filters. |
|
| Scans username presence across 26 coding, academic, and creative networks. |
|
| Precision search dorking ( |
|
4. Technical, Domain & Network Intelligence
Tool | Purpose | Key Parameters |
| Inspects domain SSL/TLS certificate validity, DNS, and RDAP/WHOIS. |
|
| Detects frontend frameworks (React, Next.js, Vue), CMS, CDN, and servers. |
|
| Public IP Geolocation, ASN, ISP, and Organization intel. |
|
| Historical time-travel & deleted webpage snapshot scraper. |
|
| Certificate Transparency subdomains discovery in <1 sec. |
|
| Deep DNS records (IPv4, IPv6, MX) infrastructure audit. |
|
๐ง Autonomous AI Agent Playbook
InfinityScrape includes an advanced Cognitive Reasoning Framework (skills/infinity-scraper/SKILL.md) that teaches autonomous AI agents how to:
Dynamically deconstruct user prompts into search and location clues.
Multi-tool chain across tools (e.g.
Search โ Filter โ Batch ScrapeorGeocode โ Locality Dork โ Profile Extraction).Auto-escalate from fast TLS to Playwright headless browser when encountering dynamic React single-page apps.
๐ Read the full agent playbook: skills/infinity-scraper/SKILL.md
๐ป Command-Line Interface (CLI)
You can also use InfinityScrape directly from your terminal:
# Scrape a URL to Markdown
python -m infinity_scraper.cli scrape "https://example.com"
# Scrape dynamic SPA with infinite scroll
python -m infinity_scraper.cli scrape "https://news.ycombinator.com" --browser --scroll 3
# Live search and auto-scrape top results
python -m infinity_scraper.cli search "Quantum computing breakthroughs" --scrape --max 4
# Crawl documentation tree
python -m infinity_scraper.cli crawl "https://docs.python.org/3/library/asyncio.html" --pages 5 --depth 2
# Extract remote PDF
python -m infinity_scraper.cli pdf "https://example.com/report.pdf" --pages 10๐งช Running Automated Tests
Run the comprehensive unit and integration test suite:
python -m tests.test_scraperTest Coverage:
โ Fast TLS Impersonator
โ Playwright Dynamic Browser
โ DuckDuckGo Live Search
โ HTML Table to Markdown Converter
โ Recursive BFS Documentation Crawler
โ Zero-GPU YouTube Transcript Extraction
โ OSINT SSL, IP Intel & 25+ Platform Presence Check
๐ License & Copyright Protection
This project is licensed under the MIT License (with Mandatory Attribution & DMCA Enforcement).
Mandatory Attribution Notice:
You are free to use, modify, and integrate this project for commercial or personal use.
However, the original Author Attribution and Copyright notice MUST be preserved in all copies, forks, or derivative distributions.
Removing the author's name/credits and re-uploading/pushing to GitHub as your own work is strictly prohibited and constitutes copyright infringement. Any infringing repository is subject to immediate GitHub DMCA Takedown & Repository Deletion and legal enforcement.
Author / Creator: VirajVerse
Repository: https://github.com/virajverse/infinity-scraper-mcp
See
LICENSEfor complete legal terms.
This server cannot be installed
Maintenance
Resources
Unclaimed servers have limited discoverability.
Looking for Admin?
If you are the server author, to access and configure the admin panel.
Related MCP Servers
- AlicenseNot gradedqualityAmaintenanceEnables AI agents to perform undetectable browser automation that bypasses Cloudflare, antibots, and social media blocks. Provides 105 tools for element extraction, network debugging, and real-world web scraping with a 98.7% success rate on protected sites.1,589MIT
- AlicenseAqualityBmaintenanceEnables AI agents to scrape any website by providing tools for JavaScript rendering, antibot bypass, and automatic captcha solving. It supports synchronous, asynchronous, and batch scraping operations with built-in proxy rotation.5207MIT

ScrapeLab MCPofficial
AlicenseNot gradedqualityDmaintenanceEnables undetectable web scraping and browser automation for AI agents with 84 tools including stealth navigation, element extraction, network interception, and auto cookie consent dismissal. Bypasses anti-bot systems like Cloudflare and DataDome while providing LLM-ready markdown output and full Chrome DevTools Protocol access.MIT
HasData MCP Serverofficial
AlicenseAqualityAmaintenanceDirect access to 40+ scraping and search tools. Extract structured data from Google (Search, Maps, Trends), Amazon, Airbnb, Social Media, and any web page directly into your AI agent.2424MIT
Related MCP Connectors
Give your agent live data from Twitter, Reddit, the web and GitHub. No API keys, no scraping stack.
Reliable web access for AI agents: smart HTTP, rotating proxies, and full-browser rendering.
Enable language models to perform advanced AI-powered web scraping with enterprise-grade reliabiliโฆ
Latest Blog Posts
- Who's Calling? MCP Hosts Are an Identity Blind Spot (And the Spec Knows It)By Om-Shree-0709 on .mcpAgent IdentityOAuth 2.1
- Your AI Chatbot Just Exposed Your CEO's Salary to an InternBy Om-Shree-0709 on .Agent IdentityMCP SecurityOAuth Delegation
- Why MCP Servers Need Execution Sandboxing (And Why Your Current Stack Isn't Enough)By Om-Shree-0709 on .Agentic AiPrompt InjectionWebAssembly
MCP directory API
We provide all the information about MCP servers via our MCP API.
curl -X GET 'https://glama.ai/api/mcp/v1/servers/virajverse/infinity-scraper'
If you have feedback or need assistance with the MCP directory API, please join our Discord server