get_crawl_issue_summary
Roll a crawl's alerts up into unique issues to fix, ranked by reach. The checks emit one alert PER PAGE, so a single dead external URL linked from 400 pages looks like 400 findings in get_crawl_alerts; here it is ONE item with target=the dead URL and pageCount=400. Prefer this over get_crawl_alerts when answering 'what should I fix first' or 'which broken links does this site have'. Items are grouped by (type, affected URL); checks about the page's own text group by that text instead, so TITLE_TOO_LONG yields one item per offending title with target=the title and targetIsSubject=true. Checks naming neither (e.g. TITLE_MISSING) yield one item per type with target=null. Narrow with category (e.g. 'External Links') or type. Ignored alerts are excluded unless includeIgnored=true. Each item's pageUrls list is capped - trust pageCount for reach and use get_crawl_alerts for the full page list.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| page | No | Zero-based page index (default 0). | |
| size | No | Groups per page (default 50, capped). | |
| type | No | Alert type to filter by, e.g. EXTERNAL_LINK_4XX, LINKS_TO_BROKEN. Takes precedence over category. | |
| crawlId | Yes | Path parameter crawlId. | |
| category | No | Catalog category to filter by, e.g. 'External Links', 'Title', 'Images'. Use get_crawl_issues to list them. | |
| severity | No | Filter by severity: ERROR, WARNING, NOTICE. Accepts a comma-separated list, e.g. 'ERROR,WARNING'. | |
| includeIgnored | No | Include ignored alerts (default false). | |
| pageUrlContains | No | Only alerts whose page URL contains this substring. |