Create a watched project
create_projectCreate a project that watches a site over time: it is crawled on its schedule and each run is compared with the last, recording pages added, modified and removed. Use it when the request is to monitor or track a site; for a one-off read use crawl_site (keep_crawl_as_project can turn that into a project later), and check list_projects first so the site is not added twice. Creating it reads the site's robots.txt and sitemaps but fetches no pages; every run then spends credits per page. Refused when the plan's project limit is reached. Returns the project, whose id start_run takes to crawl it now.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| name | No | A name for the project; omit to use the site's host name. | |
| seed | Yes | Where the crawl starts: a public site's URL or bare domain, e.g. 'https://example.com/blog/' or 'example.com'. The project covers that host. | |
| config | No | Any subset of the settings describe_project_config lists. For one section of a site pass include_paths, e.g. {"include_paths": ["/blog/*"]}; map_site shows the site's sections first. Omit for the defaults. | |
| schedule | No | How often it re-crawls on its own; 'manual' (default) runs only when start_run is called. | manual |