YandexAdditionalBot
The training scraper YandexAdditionalBot is operated by Yandex. It collects text in bulk for model training, which is a separate question from whether an AI product can cite the site today. It appears in server logs as yandexadditionalbot.
What it feeds
YandexAdditionalBot exists so that robots.txt can be applied to Yandex AI answers rather than to the search index.[1] It works over pages the main crawler has already indexed and does not request indexing itself.[2]
How to identify it
Blocking and its consequences
Disallowing the token is the documented route to keeping a page out of Yandex AI answers while leaving ordinary Yandex search indexing alone. Blocking the main crawler instead removes the site from both at once.
Verification
There is no documentation page or IP range list from Yandex that we can point at for this token.
A user agent is a claim, not a proof: any client can send any name. Where an operator publishes IP ranges, a reverse lookup is the only way to tell a real visit from an impersonation.
Seen by Baseline
Requests from this bot across sites tracked by Baseline, last 30 days. Aggregated over - sites, never reported per site.
References (4)
Last updated 2026-09-05. Written and maintained by Baseline Labs.