CleanWeb x402 — Smart Web Scraping & YouTube AI Agent
Server Configuration
Describes the environment variables required to run the server.
| Name | Required | Description | Default |
|---|---|---|---|
No arguments | |||
Instructions
Guidance the server publishes about itself, which clients place ahead of the tool catalog so the model reads it before choosing anything.
This server publishes no instructions, or was last inspected before Glama recorded them.
Capabilities
Features and capabilities supported by this server
Protocol revision2025-11-25
| Capability | Details |
|---|---|
| tools | {
"listChanged": false
} |
| prompts | {
"listChanged": false
} |
| resources | {
"subscribe": false,
"listChanged": false
} |
| experimental | {} |
Tools
Functions exposed to the LLM to take actions
| Name | Description |
|---|---|
| get_payment_infoA | Returns complete Web3 x402 micropayment configuration, supported multi-chain USDC contract addresses, EVM network chain IDs (Polygon: 137, Base: 8453, Arbitrum: 42161), recipient wallet address, and pricing tiers for all CleanWeb Studio agent tools. Usage Guidelines:
|
| clean_web_contentA | Scrapes and converts any target web page into clean, LLM-ready structured Markdown, stripping ads, cookie banners, navigation clutter, modals, and script noise. Usage Guidelines:
|
| clean_youtube_transcriptA | Extracts high-precision subtitles, timestamped transcripts, and comprehensive AI summaries for public YouTube videos using Google Gemini Flash intelligence. Usage Guidelines:
|
| clean_pdf_researchA | Parses and extracts structured plain text, sections, and academic metadata from online PDF whitepapers and research papers. Usage Guidelines:
|
| get_vault_balanceA | Checks the remaining pre-funded USDC balance, total usage, and session status for an agent wallet address or session key. Usage Guidelines:
|
| oracle_groundingA | Executes real-time web search, noise-free markdown extraction, Gemini AI JSON structuring, and cryptographically signs the result with an on-chain verifiable EIP-712 attestation (0.035 USDC). Usage Guidelines:
|
| verify_oracle_attestationA | Verifies an EIP-712 cryptographic attestation produced by CleanWeb Oracle off-chain without consuming gas. Usage Guidelines:
|
| clean_text_rawB | Extracts pure, tag-free plain text optimized for RAG embedding and vector indexing (0.001 USDC). Usage Guidelines:
|
| map_siteA | Discovers domain sitemap or traverses internal anchor links to return a canonical URL tree (0.002 USDC). Usage Guidelines:
|
| search_web_quickB | Performs fast real-time keyword web search returning titles, links, and text snippets (0.002 USDC). Usage Guidelines:
|
| extract_json_schemaA | Extracts schema-constrained structured JSON data from any webpage using Gemini AI (0.030 USDC). Usage Guidelines:
|
| deep_research_topicA | Performs multi-source web crawling and AI synthesis to generate an executive research briefing (0.150 USDC). Usage Guidelines:
|
| clean_batch_scrapeA | Concurrently scrapes and extracts clean markdown from up to 10 URLs in parallel with high-speed async processing (0.005 USDC). Usage Guidelines:
|
| get_pass_statusA | Checks the active subscription status, remaining query quota, and validity period for an agent EVM wallet address (0x...) or Agent VIP Pass Token. Usage Guidelines:
|
Prompts
Interactive templates invoked by user choice
| Name | Description |
|---|---|
No prompts | |
Resources
Contextual data attached and managed by the client
| Name | Description |
|---|---|
No resources | |
TDQS
Scored across 14 tools
Every tool targets a distinct resource and operation: separate scrapers for web, YouTube, PDF, batch, and raw text; distinct payment/status/balance tools; clear separation between quick search, deep research, and oracle-grounded search. The descriptions explicitly cross-reference other tools with 'Do NOT use' guidance, eliminating ambiguity.
Most names follow a consistent verb_noun pattern (get_payment_info, clean_web_content, map_site, verify_oracle_attestation). Minor deviations include 'oracle_grounding' (noun_noun) and 'clean_batch_scrape' (verb-adjective-noun), which are still understandable but slightly break the pattern.
14 tools is well within the ideal 3–15 range and each tool earns its place by covering a distinct aspect of web research, extraction, oracle verification, and payment management. The count feels substantial but not bloated.
The tool surface is remarkably comprehensive for its stated domain: single and batch scraping, YouTube and PDF handling, raw text extraction, quick search, deep research, site mapping, JSON schema extraction, oracle signing/verification, and payment/pass/balance queries. No obvious dead ends or missing lifecycle operations are evident.