What is GrokBot?
GrokBot is the GrokBot crawler entry attributed to xAI for gathering public web material used in training the Grok model family. xAI has published little crawler-specific documentation for this token. The operator and stated purpose are known, but the entry remains marked unverifiable because there is no recorded technical method for proving that a request came from xAI.
The name now needs care. xAI also documents a live Web Search tool for Grok that can browse pages when answering a query, and its current Grok Bot product uses persistent agents that operate websites and apps. Those products do not establish that a request carrying GrokBot is a live search fetch or an interactive agent action. The supplied agent directory identifies the newer agent with GrokAgent, a different token.
For this crawler, the documented implication is training data collection. A fetched page may contribute material to a future Grok model, but a single request cannot show whether xAI retained the content or used it in a training run. It also does not guarantee that Grok will mention, recommend, or cite the page in an answer.
That difference separates model development from real-time retrieval. A site may allow live answer-time access while declining AI training, but xAI has not published a complete token map that would let GrokBot represent both choices. Rules and log analysis should therefore stay tied to the exact user agent observed.
GrokBot is classified as only partially respecting robots.txt. A group for GrokBot still records the site's preference and may stop compliant crawling, but it should not be treated as a guaranteed access control. The available facts do not explain which requests honor rules or why behavior is partial.
There is no published IP range or Web Bot Auth directory for this crawler in the source record. Anyone can copy a user-agent string, so a GrokBot log match is a classification signal rather than cryptographic identity. If exclusion is mandatory, combine robots.txt with an edge or server policy and review the resulting requests.
