edge_blocking
Who is actually doing the blocking: for each CDN/WAF vendor, the share of (domain x crawler) pairs that robots.txt ALLOWS and the server refuses anyway — i.e. how much of the blocking is an infrastructure default rather than a decision the site owner made. Each vendor also comes broken down per crawler, which separates a blanket wall (same rate for every crawler) from a managed block list that names some AI user-agents and not others.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| limit | No | How many vendors to return, highest contradiction rate first (default 12). Vendors with fewer than 200 allowed pairs are left out of the table rather than reported on thin evidence. |