
- Top Pages: The most frequently crawled URLs on your site.
- Top Crawlers: Which AI agents are visiting most often.
- Time: When the crawl event occurred.
- Crawler: Which bot made the request.
- Status: The HTTP response code (200, 301, 404, etc.).
- Method: GET or HEAD request used by the crawler.
- Path: The specific URL accessed.
- Query String: Any parameters used in the request.
- Referrer: The source of the crawler’s request (if available).
- Status Codes – Filter by crawl result type, such as:
- 200 (Success): Successfully fetched pages.
- 301/302 (Redirects): Pages moved or temporarily redirected.
- 403 (Forbidden): Pages blocked by permissions.
- 404 (Not Found): Missing or deleted pages.
- 500 (Server Error): Failed page loads that prevent crawling.
- AI Crawler Type – Focus on specific agents like GPTBot, ClaudeBot, or PerplexityBot to analyze visibility across different AI models.
- Optimize Your Most-Crawled Pages: High crawl frequency signals authority, ensure those pages are up to date and technically healthy.
- Fix Crawl Errors Immediately: 403s, 404s, and 500s mean lost visibility. Prioritize these URLs.
- Balance Model Exposure: If GPTBot dominates crawls but ClaudeBot doesn’t appear, adjust content tone and structure for better AI diversity.
- Track Crawl Trends: Spikes in activity often follow sitemap updates or content refreshes, use them as feedback for what’s working.