Promptwatch Logo

Googlebot

Googlebot is the main crawler for Google Search.
Googlebot
Search Engine Crawler

What is Googlebot?

Googlebot is the main crawler for Google Search. It fetches pages for Google's search indexes, and rules for the Googlebot token also affect Discover and other Search features. Google publishes smartphone and desktop user agents, both of which contain the same stable token.

The crawler discovers URLs, checks the host's robots.txt, fetches allowed resources, and returns as Google decides content may need refreshing. Requests come from distributed Google infrastructure, so one site can see several source addresses. Googlebot can use ETag or Last-Modified responses when recrawling, which saves sending an unchanged page again.

Google's Search AI features reuse this Search workflow. A page must be indexed and eligible to appear in Google Search with a snippet before it can be shown as a supporting link in AI Overviews or AI Mode. Allowing Googlebot makes that possible, but it does not guarantee ranking or citation.

Googlebot access and Gemini training controls are not the same decision. Google-Extended is the separate robots token for managing certain uses of crawled content in future Gemini training and for grounding Gemini from the Search index. Blocking Google-Extended does not remove a site from Search, while blocking Googlebot can prevent fresh content from being crawled for both standard and AI Search features.

Relevant for AI searchGoogle SearchAI OverviewsGemini

Is Googlebot relevant for AI search?

Yes. Googlebot feeds Google Search, AI Overviews, Gemini, so the pages it can reach shape what those AI products say about you.

Googlebot feeds a search index that also powers that engine's AI answer features and overviews. One crawl can serve a classic results page and a generated summary, so blocking it costs you both the rankings you would expect and a growing share of AI answers built on the same index.

How to handle Googlebot

Keep public pages crawlable if they should appear in Google Search. Use path rules for faceted or otherwise duplicate URLs that Google does not need. Normal cache validators can reduce repeat transfer costs.

To stop Googlebot across the host, use its documented token:

User-agent: Googlebot
Disallow: /

Googlebot obeys robots.txt for automatic crawls. A crawl block is not a reliable removal instruction for URLs Google already knows, since an uncrawled URL can still be referenced from external links. Use the appropriate Search removal or indexing control for that separate goal. Snippet controls such as nosnippet and max-snippet also affect what Search can show in its AI features.

Examples

  • A documentation team submits a sitemap after a release, then confirms that Googlebot revisited the changed URLs before checking their Search status.
  • A publisher allows Googlebot for Search eligibility but adds a separate Google-Extended rule because its policy for Gemini-related reuse is different.

Frequently asked questions about Googlebot

Learn about AI visibility monitoring and how Promptwatch helps your brand succeed in AI search.

Google says the rules affect Google Search, including Discover and Search features, along with products that rely on its crawling such as Google Images, Video, and News.

Be the brand AI recommends

Monitor your brand's visibility across ChatGPT, Claude, Perplexity, and Gemini. Get actionable insights and create content that gets cited by AI search engines.

Promptwatch Dashboard