What is BraveBot?
BraveBot crawls pages for the Brave Search index. That index supplies conventional search results and can ground answers from Brave's Leo assistant, so this is both a search crawler and part of an AI answer workflow. It is not a user agent that appears only after one person asks Leo to open a particular URL.
Robots.txt controls whether Brave fetches a page, and BraveBot is classified as respecting crawl directives. Brave also warns that blocking a crawl is not the same as removing a URL from its index. For delisting, it directs site owners to use a noindex directive and let Brave fetch the updated page.
Brave describes its crawler as a discovery system for finding new pages and indexing their content. The company also says the crawl is partially informed by its Web Discovery Project, an optional feature in the Brave browser. A visit may therefore be part of initial discovery, a refresh of known content, or another indexing pass.
Brave's current crawler documentation has an unusual identity policy. It says the crawler does not advertise a differentiated user agent and will not crawl a page that Googlebot cannot crawl. The catalog nevertheless has a stable BraveBot token, while Cloudflare's bot directory records Bravebot/1.0 inside a full Chrome-style user agent. Cloudflare's bot directory marks that recorded identity as not following robots.txt, so this page's positive classification relies on Brave's own stated Googlebot boundary. A rule for the token is useful for traffic that presents it, but it may not describe every Brave Search request.
Pages that Brave can fetch are candidates for Brave Search and for grounded Leo responses. Blocking crawl access can limit updates and reduce the chance that current material is available on those surfaces, but a successful fetch does not guarantee a ranking, quotation, or citation. There is no evidence in the operator record that BraveBot collects this content for general model training.
Cloudflare's recorded user agent is Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; Bravebot/1.0; +https://search.brave.com/help/brave-search-crawler) Chrome/W.X.Y.Z Safari/537.36. No Web Bot Auth directory or verified IP identity is listed for BraveBot. Combined with Brave's nondifferentiated user-agent policy, the header should be treated as a classification clue rather than conclusive authentication.
