Promptwatch Logo

Cohere AI

Cohere AI collects publicly available web text that helps train and refine Cohere's large language models for enterprise generative AI.
Coherecohere-ai
AI CrawlerAI Training

What is Cohere AI?

Cohere AI is the directory entry for requests containing the token cohere-ai. The entry classifies it as an AI training crawler and retains a description that says it collects public web text for Cohere's models. Cohere's current operator documentation does not confirm that identity.

In its web crawler policy, Cohere says it does not currently use Cohere bots or user agents to crawl or scrape web content for generative AI foundation-model training. Its table of active crawler names contains no entries. The policy says Cohere would identify any future training crawlers and publish blocking instructions.

That conflict changes how a log match should be interpreted. A cohere-ai header is not current evidence that Cohere fetched the page or placed its content in a training dataset. The available sources do not establish whether the token is obsolete, used by another client, or simply spoofed, so none of those explanations should be asserted without request-level evidence.

Cohere models can be connected to search tools by developers, but Cohere's documentation has the external tool perform the search. Traffic from such a setup would carry the identity chosen by that tool or application. The cohere-ai token is not documented as Cohere's public AI search agent, and allowing it has no verified effect on citations or search visibility.

This entry marks robots.txt support as true. Cohere's current policy says any future crawler must be designed to respect robots.txt, but its example uses the hypothetical token Coherebot, not cohere-ai. Robots behavior for requests using this page's token is therefore unconfirmed despite the retained metadata value.

There is no Cloudflare bot-directory record, Web Bot Auth signature directory, or verified IP status for cohere-ai in the supplied facts. A site can classify or block the literal token, but should not grant access or attribute the request to Cohere on that basis alone.

Relevant for AI searchCohere

Is Cohere AI relevant for AI search?

Yes. Cohere AI feeds Cohere, so the pages it can reach shape what those AI products say about you.

Cohere AI gathers public web content that can end up in the training data for large language models. Once your pages are in that set, they influence how the operator's models talk about you for that model generation. Allowing it lets your own writing carry weight in those answers; blocking it means the models learn about you from third parties instead.

Track Cohere AI with Promptwatch

Promptwatch classifies Cohere AI (cohere-ai) in real time from your server and CDN logs. See exactly when it visits, which pages it requests, the status codes it gets, and how those crawls map to AI citations.

How to handle Cohere AI

Do not allowlist cohere-ai as trusted Cohere traffic. Cohere currently lists no active training crawler, and this identity has no published signature or IP verification method.

You can still state a policy for the recorded token:

User-agent: cohere-ai
Disallow: /

Treat that rule as a preference for clients that choose to honor robots.txt, not proof of a Cohere opt-out. If requests continue or the content must remain inaccessible, block the token at the edge and keep the underlying pages behind authentication. Recheck Cohere's official crawler table before changing the policy for any newly announced bot name.

Examples

  • A security team sees `cohere-ai` in an access log, finds no matching bot in Cohere's current crawler table, and records the source as unverified instead of labeling the request as Cohere training traffic.
  • A publisher blocks the literal token at its edge but keeps watching Cohere's official policy for a future named crawler with documented robots.txt instructions.

Frequently asked questions about Cohere AI

Learn about AI visibility monitoring and how Promptwatch helps your brand succeed in AI search.

No. Cohere's current crawler policy lists no active bot or user agent for foundation-model training.

Be the brand AI recommends

Monitor your brand's visibility across ChatGPT, Claude, Perplexity, and Gemini. Get actionable insights and create content that gets cited by AI search engines.

Promptwatch Dashboard