ScrapeMole Reader: web pages to Markdown, link previews, RSS feeds, links and sitemaps for AI agents: Sitemap reader
sitemapRead-onlyIdempotent
Use when you need the list of pages a site publishes (to crawl it, find recent posts, or audit SEO): give a site URL (the sitemap is discovered from robots.txt Sitemap: lines, then /sitemap.xml, /sitemap_index.xml, /wp-sitemap.xml) or a sitemap URL directly. Returns the page URLs with lastmod, changefreq and priority; for a sitemap index, returns its child sitemaps and (expand=true, default) the URLs of the first child sitemaps up to the limit. Gzipped sitemaps supported. Paid tool: x402 USDC on Base, $0.002 per successful call; call without payment to receive the PaymentRequired object.
Input Schema
TableJSON Schema
| Name | Required | Description | Default |
|---|---|---|---|
| url | Yes | Site URL (e.g. https://example.com) or a sitemap .xml/.xml.gz URL. | |
| limit | No | Max page URLs returned (1-5000). | 1000 |
| expand | No | For a sitemap index, also read the first child sitemaps (true/false). | true |