scrapy_cloud_run_script
Start a Scrapy Cloud job that runs a standalone Python script deployed with the project (declared under 'scripts' in its setup.py), instead of a spider. Pass the script's command-line arguments as one string in 'args'. Returns the job key and dashboard URL; the job starts in state 'pending'. Check progress with scrapy_cloud_get_job and read its output with scrapy_cloud_get_job_log.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| args | No | Command-line arguments for the script, as one string, e.g. '--limit 100 --dry-run'. | |
| tags | No | Tags to add to the job. | |
| units | No | Scrapy Cloud units for the job. Default is the project's setting. | |
| script | Yes | Script file name as deployed, e.g. hello.py; the 'py:' prefix Scrapy Cloud uses is optional. | |
| priority | No | Queue priority, 0 (lowest) to 4 (highest). Default 2. | |
| project_id | Yes | Scrapy Cloud project id (the numeric id in the dashboard URL). | |
| job_settings | No | Scrapy settings overriding the project's, readable by the script via sh_scrapy.utils.get_project_settings. |