Skip to main content
Glama
aidvizhhub

camoufox-research

by aidvizhhub

Пакетное чтение

batch_fetch
Read-onlyIdempotent

Fetch 10-50 URLs in one browser call for deep research, returning article text with caching, rate limits, and parallel workers to avoid cold starts.

Instructions

Открывает НЕСКОЛЬКО URL в одном браузере — для глубокого ресёрча на 30-50 источников одним вызовом вместо серии холодных стартов. Кэш: уже посещённые URL возвращаются мгновенно, без браузера. Rate limit между переходами защищает от капчи. Батч ≥8 URL — параллельно (пул потоков, свой браузер на поток); число воркеров автоопределяется по ресурсам машины (слабый ПК — 1-2, мощный — 3-4), max_parallel — явное ограничение. Возвращает тексты с разделителями '--- URL: ...'. article_only=True — извлечь текст статьи (Trafilatura), без меню и баннеров. Пример: batch_fetch(urls=["https://docs.python.org/3/", "https://opencode.ai/docs/"], max_chars=20000, article_only=True) ХОЧУ ПОЛНОТУ → max_chars=20000..100000 + article_only=True + max_parallel=4; экономия контекста → max_chars=4000 при 30+ URL. Замер 21.09 (8 URL разных доменов, холодный кэш): 4000 → 32 598 симв. за 36.4с (все 8 обрезаны ровно по 4000), 12000 → 96 598 симв. за 30.3с, 20000 → 154 449 симв. за 45.9с — время НЕ растёт от объёма: страница и так читается до потолка кэша (100k), max_chars режет только ОТВЕТ. КОГДА: читать 10-50 URL одним вызовом (глубокий ресёрч после research(queries=[...]) или crawl/map_site). НЕ КОГДА: 1-2 страницы → fetch_page; URL ещё не собраны → research, sitemap, map_site сначала.

Input Schema

TableJSON Schema
NameRequiredDescriptionDefault
urlsYes
max_charsNo
article_onlyNo
max_parallelNo

Output Schema

TableJSON Schema
NameRequiredDescriptionDefault
resultYes

Schema Changelog

Changes observed during successful MCP inspections.

  1. Changed1 schema field changedv0.21.1
    • changedInput schema / properties / max_chars / default
      Previous value: -4000New value: +12000
  2. First observedv0.1.0

TDQS

A4.9/5.0
Behavior5/5

Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?

Annotations cover the safety profile (readOnly, idempotent, non-destructive, openWorld), and the description adds substantial non-structured behavior: cache hits return instantly without a browser, inter-navigation rate limiting guards against CAPTCHA, batches of ≥8 URLs run in a thread pool with per-thread browsers, and max_chars truncates only the response, not the crawl. This is richer than typical annotation coverage.

Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.

Conciseness4/5

Is the description appropriately sized, front-loaded, and free of redundancy?

Dense but front-loaded: purpose, then cache/rate-limit/parallel behavior, then an example, then when/not-when. The dated benchmark timings (36.4s/30.3s/45.9s) are somewhat verbose, but they justify the tuning advice and are not pure filler.

Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.

Completeness5/5

Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?

An output schema exists, so return values need not be explained, yet it still notes the '--- URL: ...' text separators. Given four parameters, no schema descriptions, and a parallel/cached execution model, the description supplies everything an agent needs to call it correctly.

Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.

Parameters5/5

Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?

Schema description coverage is 0%, so the description must carry all four parameters — and it does: urls (scale guidance), max_chars (concrete ranges 4000-100000 with measured tradeoffs), article_only (Trafilatura extraction excluding menus/banners), and max_parallel (explicit cap on auto-detected workers). A worked example demonstrates combined usage.

Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.

Purpose5/5

Does the description clearly state what the tool does and how it differs from similar tools?

States a specific verb and resource ('Открывает НЕСКОЛЬКО URL в одном браузере') plus the scale it targets (30-50 sources, one call instead of cold starts). An agent can immediately distinguish it from fetch_page and the other fetch-adjacent siblings.

Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.

Usage Guidelines5/5

Does the description explain when to use this tool, when not to, or what alternatives exist?

Contains explicit 'КОГДА' and 'НЕ КОГДА' sections: use for 10-50 URLs after research/crawl/map_site; use fetch_page for 1-2 pages; gather URLs first via research, sitemap, map_site. Alternatives and exclusion conditions are named outright.

Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.