extract_from_html
Extract main content and metadata from pasted or pre-loaded HTML without network requests. Parses blocks and applies multi-strategy extraction to return markdown or text.
Instructions
Extract main content + metadata from HTML you already have (no network).
Useful when the user pasted HTML, or another tool (a browser) already loaded the page. Runs the same block detection and multi-strategy extraction as fetch_page.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| url | No | ||
| html | Yes | ||
| format | No | markdown | |
| adapter | No | auto | |
| max_chars | No |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
No arguments | |||