audit_crawler_access
Test whether AI crawlers such as GPTBot and ClaudeBot can fetch a URL. Parse robots.txt and run live GETs per bot User-Agent to detect access blocks.
Instructions
Verify that major AI crawlers (GPTBot, OAI-SearchBot, ClaudeBot, PerplexityBot, CCBot, Google-Extended, Applebot-Extended, Bytespider, Meta-ExternalAgent, plus real-time fetch UAs) can fetch a URL. Parses robots.txt and does a live GET with each bot's User-Agent. Surfaces robots.txt blocks AND UA-based gating that breaks AI citation.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| url | Yes | Page URL to test for AI crawler access. | |
| bots | No | Override the default bot list. Each entry is a User-Agent token (e.g. 'GPTBot', 'ClaudeBot'). | |
| fetch_with_ua | No | If true, do a live GET as each bot's User-Agent and report status. Disable to only parse robots.txt (no extra requests). |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
| url | Yes | The URL that was audited. | |
| bots | Yes | Per-bot access verdict combining robots.txt + live UA test. | |
| note | No | ||
| summary | Yes | ||
| fetched_at | Yes | UTC ISO-8601 timestamp. | |
| robots_url | Yes | robots.txt URL that was parsed. | |
| robots_error | Yes | Error message if robots.txt fetch failed. | |
| robots_status | Yes | HTTP status of the robots.txt fetch. | |
| robots_present | Yes | Whether a non-empty robots.txt was found. |