Check a crawl
writ_crawl_statusRead-onlyIdempotent
Status of a crawl by id: page counts, status, the dataset workflow id. With wait=true one call holds up to 75 s and returns when the crawl converges, with the collected rows inline (data, shaped by output); it resolves a crawl tool's 504 / crawl-id handle. A crawl still running at the ceiling answers with its current status, and the same call can be repeated.
Input Schema
TableJSON Schema
| Name | Required | Description | Default |
|---|---|---|---|
| wait | No | Hold until the crawl is terminal and inline its rows (default false). | |
| limit | No | Rows to inline when it converged (default 50). | |
| output | No | Response shape, for an answer a program or an API consumes rather than a reader. {shape: 'envelope' (default: Writ's full answer, projected) | 'table' ({columns, rows, total}) | 'records' (bare list of records) | 'record' (the newest record alone: one entity, a usage meter, a dashboard), fields: ['used', 'percent_used as pct', 'items.0.price as first_price'] (ordered pick, renames, dotted paths; a missing path is null, so keys are stable), exclude: ['depth'], include_meta: false (page metadata content_kind/depth/thumbnails is stripped unless true), key: 'usage' (wrap)}. On writ_crawl_site with save_as it is saved as the API's default shape. | |
| crawl_id | Yes | Crawl id from writ_crawl_site. | |
| preview_chars | No | Cut inline text cells to this many chars (default 12000; 0 = full). | |
| timeout_seconds | No | Ceiling for wait=true (≤75). |