What is Yandexbot?
YandexBot is the main indexing robot for Yandex Search. It downloads pages for Yandex's search database under a user agent containing YandexBot/3.0 and the operator's yandex.com/bots reference URL.
A visit means the page is being considered or refreshed for the ordinary Yandex index. Access does not guarantee inclusion or a particular ranking. Blocking the main robot prevents it from reading the affected URLs, which is relevant to Yandex Search but has no direct effect on indexes run by other companies.
Yandex operates many specialized robots, and their rules differ. The operator's robot identity table marks the main YandexBot as following general robots.txt rules. Other agents used for ads, screenshots, or user-triggered functions may ignore broad wildcard restrictions and need their own user-agent records.
Yandex also documents YandexAdditional and YandexAdditionalBot separately. Those agents help apply robots controls so content already indexed by the primary crawler does not appear in Search with Yandex AI responses; they do not make indexing requests. This means a YandexBot rule is a search indexing control, not a precise AI-only setting. The supplied record does not establish that YandexBot collects pages to train a generative model.
The browser version appended to the full user agent can change, so it is a poor filter. Yandex recommends reverse DNS for verification: the hostname must end in yandex.ru, yandex.net, or yandex.com, and a forward lookup must return the original address. Its robots use changing addresses in AS13238, AS208722, and AS212066 rather than a fixed published list.
Cloudflare's bot directory marks its Yandex identity as not following robots.txt, but its pattern is the shared http://yandex.com/bots URL found in many different Yandex user agents. That broad match does not override Yandex's token-specific statement for the main indexer. Cloudflare supplies no signed-agent key directory for this bot.
