Download web pages to files
downloadFetch multiple URLs and save each as Markdown, HTML, PDF, screenshot, MHTML, or raw source files, returning saved paths. Choose format and options like fit, citations, and timeouts.
Instructions
Fetch one or more URLs and save each result as a file in a directory on the server, returning the saved paths. Supports Markdown (default), HTML, PDF, PNG screenshot, MHTML, and the raw source. PDFs are transcribed to Markdown unless format is 'raw'; non-web content (images, archives, ...) is saved as served. Existing files are overwritten. The call fails only if no file was saved.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| fit | No | Keep only the main content (drops menus, footers, and the like); falls back to the full page when nothing is left. Markdown only. | |
| urls | Yes | URLs to fetch (http or https), at most the server's per-call limit (see the server instructions); duplicates are fetched once. | |
| format | No | File format: 'markdown' (default), 'html', 'pdf' (page printed to PDF), 'screenshot' (PNG), 'mhtml' (single-file web archive), or 'raw' (the original response bytes, e.g. a PDF or image as served). | markdown |
| citations | No | Turn links into numbered references listed at the end. Markdown only. | |
| directory | No | Subdirectory of the server's download directory to save into; must stay inside it. Defaults to the download directory itself. | |
| timeout_s | No | Page load timeout per URL in seconds; defaults to the server setting. | |
| ignore_links | No | Drop links from the Markdown output. Markdown only. | |
| ignore_images | No | Drop images from the Markdown output. Markdown only. |