List a crawl's captured files
writ_crawl_filesRead-onlyIdempotent
The original documents a crawl captured (PDFs, office docs, images, CSVs) as stored files: filename, size, version, source_url, and a short-TTL download_url fetchable with no further auth. The crawl's dataset holds the extracted text; this returns the files themselves. crawl_id selects one run; crawl (saved crawl slug/name/id) its most recent completed run(s).
Input Schema
TableJSON Schema
| Name | Required | Description | Default |
|---|---|---|---|
| runs | No | With `crawl`: how many recent completed runs to aggregate (default 1 — the current version of every document). | |
| crawl | No | Saved crawl slug, name, or id (alternative to crawl_id). | |
| limit | No | Max files to return (default 100, cap 200). | |
| crawl_id | No | Crawl id from writ_crawl_site. |