Promptwatch Logo

Coveobot

Coveobot is a crawler operated by Coveo that indexes content for enterprise search, recommendations, and generative experience platforms.
UnverifiableCoveoCoveobot
AI AssistantSearch Engine Crawler

What is Coveobot?

Coveobot is the crawler Coveo uses when a customer configures web content for a Coveo index. Coveo's platform combines material from websites and other repositories so that the customer can offer search, recommendations, customer-service experiences, or generative answers. The crawler is therefore tied to an enterprise content source rather than a universal consumer search engine.

The resulting index can support a generative experience, so pages available to Coveobot may influence answers inside that customer's Coveo implementation. This is not the same as broad visibility across public AI assistants. Coveo's crawler documentation does not state that a page fetched by Coveobot is used to train a general-purpose model.

For a Coveo Web source, an administrator supplies a starting URL and the crawler discovers pages through site navigation and links. The source configuration controls what content is indexed and how it is retrieved. Coveo can also retain source permissions so that indexed items remain visible only to users who are authorized to see them.

The standard header recorded by Coveo and Cloudflare is Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko) (compatible; Coveobot/2.0;+http://www.coveo.com/bot.html). Coveo's default robots matching name is CoveoBot. There is no Web Bot Auth key directory or verified IP identity in this bot, so the user-agent text cannot authenticate a request and the bot remains classified as unverifiable.

A Coveobot visit usually means that a particular organization chose the site as one of its content sources. If the site belongs to that organization, the requests may be refreshing its own support, commerce, or documentation index. If another party configured the URL, the destination may not have a direct relationship with the Coveo customer.

Coveo's Web source respects robots.txt page restrictions and Crawl-delay by default, which is why this bot is classified as respecting robots.txt. Administrators can enable an override that tells a Web source to ignore those directives, and Cloudflare's directory labels the agent as not following robots.txt. A rule is meaningful for the default setup, but it cannot guarantee the behavior of every customer-configured source.

Relevant for AI search

Is Coveobot relevant for AI search?

Yes. Coveobot collects pages for an AI product, so what it can crawl influences how AI systems describe your brand.

Coveobot crawls and indexes pages so the AI search or assistant behind it can retrieve them at answer time. A page it has never fetched cannot be quoted, summarized, or linked in that product's answers, so most sites keep it allowed to stay citable. Blocking it removes your pages from that AI surface and hands those citations to competitors.

How to handle Coveobot

Allow Coveobot when your organization intentionally uses Coveo to index the site. Review source scope in the Coveo administration console so the crawler sees only the intended public or permission-aware content. For unrelated crawl traffic, decide access according to the site's content policy rather than assuming a public search benefit.

Coveo's default Web source should apply this directive:

User-agent: Coveobot
Disallow: /

A Coveo customer can configure the source to ignore robots.txt or replace its user agent. Watch for continued access after publishing the rule. Since the standard header has no signature or verified IP binding, protect nonpublic content with authentication instead of a user-agent allowlist.

Examples

  • A support organization configures its help center as a Coveo Web source so new articles become searchable in its customer portal.
  • A site administrator disallows a staging directory and confirms that the default Coveo source stops requesting those URLs.
  • A security team refuses to treat `Coveobot/2.0` as authenticated because the same header can be sent by an unrelated client.

Frequently asked questions about Coveobot

Learn about AI visibility monitoring and how Promptwatch helps your brand succeed in AI search.

A Coveo customer configures a content source and its starting URL. Coveobot then retrieves pages for that customer's Coveo index rather than choosing sites for a single global index.

Be the brand AI recommends

Monitor your brand's visibility across ChatGPT, Claude, Perplexity, and Gemini. Get actionable insights and create content that gets cited by AI search engines.

Promptwatch Dashboard