hsh-web-scrape
Extract structured records from a public web page. Finds the repeating block that holds the records, pulls the fields you name out of each one, follows pagination, removes duplicates, and reports how often each requested field was actually present. Reads the HTML a site serves: pages that build their content in the browser, and anything behind a login, are refused before any charge rather than returned empty. Priced per record.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| fields | Yes | List of fields to extract per record. | |
| quantity | No | How many records to extract: a whole number 1-2,000, default 100. The price is $0.002 a record asked for, $0.02 at least; more than 2,000 is refused before payment, not cut down. | |
| source_url | Yes | Target site or section to scrape. | |
| complexity_hint | No | 'static' is the only supported value. 'js_render' and 'auth_required' are refused before any charge. |