π οΈ How-to guide Β· Last updated: September 17, 2026
<aside> πΈοΈ
WHAT IT DOES
A website crawl reads pages from a public site or help center and ingests them as content β then keeps them current on a schedule. This page covers the full setup; for monitoring and managing a crawl afterwards, see Crawls & Syncs.
</aside>
<aside> β TIP
Start narrow β depth 3 and a Daily refresh β and widen only if you're missing pages. It keeps the crawl fast and your content focused on what residents actually ask about.
</aside>
site tag is filled in for you).
The website-crawl setup β a starting URL, Max depth, Max links per page (0 = all links), and the new Max pages (0 = no limit), with more under Advanced.
Basic
| Setting | Default | Details |
|---|---|---|
| Crawl name | β | Required. A label for this crawl. |
| Starting URL | β | Required. Must begin with http:// or https://. |
| Max depth | 3 | How many link-levels deep to follow (up to 10). |
| Max links per page | 10 | How many links to follow on each page. Set 0 to follow all links (up to 100). |
| Max pages | 0 (no limit) | Total pages the crawl visits in a run. 0 means no limit β a platform ceiling still backstops it. |
| Refresh frequency | One-time | One-time, Hourly, Daily, Weekly, Monthly, or a Custom cron. Daily/Weekly/Monthly let you pick a time of day. |
| Default tags | site: <name> | Applied to every page the crawl ingests. |
Advanced