fetch_page
Retrieve public web pages as clean markdown or text with metadata, escalating from HTTP to a browser only when needed and returning clear failure reasons.
Instructions
Fetch a public web page and return clean content plus metadata.
Escalation: site adapter API -> HTTP with realistic headers -> client/UA rotation -> AMP/RSS alternates -> headless browser (if installed). Every rung is listed in fetch_attempts. On failure ok=false and blocked_reason explains why (e.g. cloudflare_challenge, captcha, verification_wall, login_required, paywall, rate_limited, not_found, soft_404, js_required).
Args: url: page URL (http/https). format: which content field(s) to return. use_browser: auto (only when cheaper rungs fail), never, or always. adapter: auto, none, wechat, github, x. max_chars: truncate content to this many characters (0 = no limit). min_chars: extracted length that counts as success.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| url | Yes | ||
| format | No | markdown | |
| adapter | No | auto | |
| max_chars | No | ||
| min_chars | No | ||
| use_browser | No | auto |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
No arguments | |||