Skip to main content
Glama
Br0ski777

web-scraper

by Br0ski777

Web Scraper to Markdown API

MCP Server x402 License: MIT

Extract clean markdown from any URL. Strips nav/ads/scripts, returns structured content. Built for RAG pipelines and AI research agents. Pay-per-call via x402 (USDC on Base L2) -- no API key, no signup, no rate-limit wall.

Part of the klymax402 marketplace -- 100 x402 micropayment APIs for AI agents, one wallet, USDC on Base.

Quickstart -- MCP

Add to your MCP client config (Claude Desktop, Cursor, ElizaOS, etc.):

{
  "mcpServers": {
    "web-scraper": {
      "url": "https://web-scraper.api.klymax402.com/mcp"
    }
  }
}

Related MCP server: x402-md-mcp

Quickstart -- HTTP (x402)

curl "https://web-scraper.api.klymax402.com/api/scrape?url=https://example.com"
# -> 402 Payment Required, with an x402 payment challenge in the response body

Any x402-aware client (@x402/fetch, x402-agent-tools, ATXP) handles the 402 -> sign -> retry cycle automatically.

Tools

Tool

Method

Path

Price

Description

web_scrape_to_markdown

GET

/api/scrape

$0.012

Scrape a URL and convert to clean markdown

web_scrape_batch

POST

/api/scrape/batch

$0.05

Scrape up to 10 URLs in batch

web_scrape_to_markdown

Scrape and extract content from a URL with full JS rendering, returned as clean markdown. Alternative to Firecrawl scrape at 2.5x lower cost. Strips navigation, ads, scripts, and boilerplate — ideal for RAG pipelines and AI research agents.

Parameters

Name

Type

Required

Description

url

string

yes

URL to scrape (e.g. https://example.com/article)

Example response:

{"title":"How to Scale APIs","description":"A guide to...","content":"# How to Scale APIs\n\nScaling requires...","wordCount":1250,"charCount":7800,"url":"https://blog.example.com/scale-apis"}

When to use: summarizing articles, building RAG corpora, researching topics from web sources, or extracting data from documentation pages. Essential for any workflow that needs to scrape and extract content from web pages as LLM input. Drop-in replacement for Firecrawl scrape.

Not for: screenshots (use capture_screenshot), SEO audit (use seo_audit_page), tech stack detection (use website_detect_tech_stack), web search (use web_search_query).

web_scrape_batch

Use this when you need to extract clean content from multiple web pages at once (up to 10 URLs). Returns the same structured markdown output as web_scrape_to_markdown for each URL.

Parameters

Name

Type

Required

Description

urls

array

yes

Array of URLs (max 10)

Example response:

{"results":[{"url":"https://a.com","title":"Page A","wordCount":800},{"url":"https://b.com","title":"Page B","wordCount":1200}],"summary":{"total":2,"totalWords":2000,"failed":0}}

When to use: building research corpora, comparing content across competitor pages, or bulk documentation extraction. Essential when you have 3+ URLs to process in one workflow.

Not for: single URLs (use web_scrape_to_markdown), SEO comparison (use seo_audit_batch).

Example agent prompts

  • "Scrape and extract content from a URL with full JS rendering, returned as clean markdown"

  • "Extract clean content from multiple web pages at once (up to 10 URLs)"

Payment

  • Protocol: x402 -- HTTP-native pay-per-call, no signup, no API key

  • Network: Base L2 (eip155:8453)

  • Asset: USDC

  • Facilitator: Coinbase CDP (primary), PayAI (fallback)

  • Also reachable via ATXP (OAuth-wrapped x402, RFC 9728 protected-resource metadata)

Part of klymax402

100 x402 micropayment APIs for AI agents -- one wallet, USDC on Base, zero signup.

License

MIT

Related MCP Connectors

Related MCP Servers

  • A
    license
    A
    quality
    B
    maintenance
    Pay-per-use clean web reader for AI agents. URL in, markdown plus metadata out, in milliseconds. Settled per-call in USDC over x402 — no signup, no API keys.
    1
    82 npm
    2
    MIT
  • A
    license
    Not graded
    quality
    D
    maintenance
    Enables extracting clean Markdown from any webpage by paying $0.005 USDC per call via the x402 protocol, with automatic wallet-based payment settlement.
    4 npm
    MIT
  • F
    license
    Not graded
    quality
    C
    maintenance
    Scrapes dynamic web pages via headless Chrome, stripping clutter and ads, and returns clean Markdown or JSON content. It integrates with Cloudflare Workflows and monetizes each extraction through the x402 protocol.
    -
  • F
    license
    Not graded
    quality
    B
    maintenance
    Enables LLM agents to convert bloated webpages into clean token-dense plain text, stripping HTML, CSS, scripts, and layout elements while reducing context-window costs and supporting x402 micro-pay-based autonomous requests.
    -