Crawl a website into the knowledge base
crawl_company_websiteIngest your company's website content into Mira knowledge base by crawling sitemap and same-origin links, extracting, chunking, and embedding readable text. Monitor crawl progress via status reads.
Instructions
Start a background crawl of a website into the company's Mira knowledge base, covering its sitemap and same-origin links while extracting, chunking and embedding readable text. Crawl progress remains available through status reads. Re-crawling an unchanged site ingests 0 new pages and is still successful. Requires the file-uploads entitlement (Growth and up).
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| url | Yes | Start URL, e.g. the site root. Same-origin pages only. | |
| language | No | Keep only pages in this 2-letter language, e.g. "de" | |
| maxPages | No |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
| jobId | No | ||
| status | No | ||
| message | No |