Running a crawl
Crawls run from the Deep Scan tab of the Sitemap Workbench. Deep Scan is available on Starter and above, to Editors and Admins.
- Open the Site, go to Sitemaps and select the Deep Scan tab.
- Click Start crawl.
- Progress updates every two seconds: pages crawled, current depth and broken links found so far.
- Click Cancel to stop early. Pages crawled up to that point are kept and analysed.
Crawler behaviour
- Flowpane identifies itself as
Flowpane/1.0 (+https://crawler.flowpane.com). - It respects robots.txt for that user agent and honours
Crawl-delay, capped at 10 seconds. - Crawls start from the homepage and follow internal links within page, depth and execution limits. A bounded crawl may stop before exploring every reachable page.
Rate limits
- One crawl per Site per hour.
- A small number of crawls run across Flowpane at once. If capacity is full, check the status shown in Flowpane; no start time is guaranteed.
See Crawl settings and limits for the per-plan limits.