scrapy_cloud_run_spider
Start a new Scrapy Cloud job for a spider now. Returns the job key and dashboard URL; the job starts in state 'pending' and is picked up by the queue. Fails with 'already scheduled' if an identical job is pending or running. Check progress with scrapy_cloud_get_job. For a recurring schedule use scrapy_cloud_create_periodic_job instead; for a standalone script use scrapy_cloud_run_script.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| tags | No | Tags to add to the job. | |
| units | No | Scrapy Cloud units for the job. Default is the project's setting. | |
| spider | Yes | Spider name, as listed by scrapy_cloud_list_spiders. | |
| job_args | No | Spider arguments, passed as -a name=value; numbers and booleans are sent as strings ('100', 'true'). | |
| priority | No | Queue priority, 0 (lowest) to 4 (highest). Default 2. | |
| project_id | Yes | Scrapy Cloud project id (the numeric id in the dashboard URL). | |
| job_settings | No | Scrapy settings overriding the project's, for example {"CLOSESPIDER_ITEMCOUNT": 100}. |