What is aiHitBot?
aiHitBot collects public information from company websites for aiHit's company-data system. The crawler's own notice says it looks for corporate information across mainly English-speaking markets and identifies its requests with the aiHitBot token.
The associated aiHitdata product builds structured company profiles from the web. Its documented fields include company descriptions, locations, contact details, registration numbers, executives, clients, partners, and investors. This is business-information extraction rather than a general web search index.
aiHitdata revisits company sites and compares newer information with earlier records. When a company changes an address, leadership listing, or another extracted field, the system can store that change as part of the company's history. Repeated aiHitBot visits may therefore be refreshes of an existing profile.
The crawler notice says it is interested in company information and avoids personal sites, personal blogs, and directories. It also says information found to be unrelated to a company is discarded. Those statements describe aiHit's collection policy, not a technical access restriction imposed by the site being visited.
There is no evidence in the supplied records or aiHit's crawler description that aiHitBot powers generative search citations or gathers a corpus for a general-purpose language model. Its documented output is a company-information database. Although this directory includes the AI Training category, blocking this token should not be presented as a proven model-training opt-out.
The bot is recorded as respecting robots.txt, and aiHit's notice makes the same commitment. It is still marked unverifiable: no Cloudflare identity record, signed-agent key directory, or verified IP range is supplied. The user-agent token supports a robots rule and log classification, but it does not authenticate the sender.
