What is Siteimprove Crawl?
Siteimprove Crawl scans sites connected to Siteimprove's content suite. It requests pages over the standard web ports, 80 for HTTP and 443 for HTTPS, and supplies crawl data for quality assurance, accessibility, policy, and SEO checks.
A Siteimprove account can run a full scan on a schedule or recheck selected pages. Siteimprove separates fetching from later analysis, so Crawler Management can show a crawl as finished while link checks and other processing are still pending. Exclusions and deduplication can also make the final product totals smaller than the raw page and link counts.
This is recurring audit traffic, not a search engine visit. Siteimprove says its crawler normally pauses between requests and can slow down when a server struggles. Customers can also exclude sections or change scan timing. Those account controls are more precise than blocking every request after it reaches the site.
The available sources document Siteimprove Crawl as a site-audit client, not a public AI answer engine, and do not connect it to model training. Fixes prompted by a Siteimprove report may improve the site itself, but allowing this crawler has no documented direct effect on AI citations or rankings.
