Crawl the site and audit every page
site_crawlCrawl a website to identify technical SEO issues: broken links, redirect chains, duplicate titles, missing H1s, noindex, thin pages, and orphans. Get actionable audit results per page.
Instructions
Breadth-first crawl of one host (respects robots.txt), auditing each page: status counts, broken links with referrers, redirect chains, duplicate titles/descriptions, missing title/description/H1, noindex, thin pages, images without alt, orphan pages, click depth and inbound links. 200 pages take 1-3 minutes.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| maxPages | No | ||
| startUrl | Yes | ||
| pathPrefix | No | Only crawl URLs whose path starts with this, e.g. '/blog/'. | |
| concurrency | No | ||
| includePages | No | Include the per-page audit rows in the response (large). | |
| includeSitemap | No | Also read the sitemap to detect orphan pages and seed the queue. |