fetch_links
Extract all outbound links from a webpage, resolved to absolute URLs with anchor text, rel/title, and internal/external classification. Skips non-web links and respects base href.
Instructions
Extract every outbound link from an HTML page, resolved to absolute URLs. Each entry includes href, anchor text, optional rel/title, and an internal/external classification (bare-domain and www. treated as the same host). Anchors (#), javascript:, mailto:, tel:, data:, and file: URIs are skipped. Respects .
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| url | Yes | ||
| limit | No | Cap on returned links (default 1000) | |
| dedupe | No | Drop duplicate hrefs (default true) | |
| filter | No | Filter by type (default 'all') | |
| max_bytes | No | ||
| timeout_ms | No | ||
| user_agent | No | ||
| max_redirects | No | ||
| allow_private_hosts | No | Allow loopback / private / link-local targets for this call (default false). Refused unless the server operator launched fetch-mcp with FETCH_MCP_ALLOW_PRIVATE_HOSTS=1 -- SSRF protection stays on by default either way. |