What is Brightbot?
Brightbot is Bright Data's named pipeline for collecting public web data across its products and services. That is broader than a search crawler with one index and one destination. Bright Data offers data tools for uses that include AI training, but a Brightbot request does not reveal the customer, dataset, or eventual use of the page.
The crawler has a cache intended to prevent another download of the same data within 24 hours, unless Bright Data approves an exception for a business reason. Bright Data also opens a health monitor for targeted domains. If its traffic correlates with slower responses, the system applies a rate limit based on the last rate that did not hurt site performance.
Bright Data publishes two identifiers that should be checked together: the user agent Brightbot 1.0 and source addresses in 82.97.199.0/24. This makes recognized Brightbot traffic easier to separate from ordinary visitors. It does not identify every request made through Bright Data's products.
Verified site owners can use the Webmaster Console to submit collectors.txt rules for personal information, private or copyrighted material, and interactive endpoints. Bright Data reviews the file before Brightbot enforces it. The company says this identity and the file currently cover Web Unlocker traffic, not Browser API or browser-based Data Collector jobs.
