List a crawl's captured files
writ_crawl_filesRetrieve the original documents a crawl captured—PDFs, office files, images, CSVs—with download URLs for when you need the actual files, not just extracted text.
Instructions
The ORIGINAL documents a crawl captured (PDFs, office docs, images, CSVs) as stored files — filename, size, version, source_url, and a short-TTL download_url fetchable with no further auth. The crawl's DATASET holds the extracted text; use this when you want the actual files. Pass crawl_id for one run, or crawl (saved crawl slug/name/id) for its most recent completed run(s).
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| runs | No | With `crawl`: how many recent completed runs to aggregate (default 1 — the current version of every document). | |
| crawl | No | Saved crawl slug, name, or id (alternative to crawl_id). | |
| limit | No | Max files to return (default 100, cap 200). | |
| crawl_id | No | Crawl id from writ_crawl_site. |