What is Alphalens Bot?
Alphalens Bot builds an index of companies and the products or services they offer. Alphalens says it wants business discovery to reflect more than a short company description, so the crawler reads company websites for the detail needed by semantic search. Its job is commercial indexing, not general web archiving.
Before fetching pages, the crawler checks the site's robots.txt file and sitemap. When a sitemap is available, Alphalens begins with its listed URLs and considers the supplied priority values. Without one, it starts at the homepage and follows links in breadth-first order to discover other pages.
The operator describes a restrained retrieval process. Alphalens Bot fetches approved HTML pages but does not submit forms, click buttons, solve CAPTCHAs, or try to pass security controls. It says it honors rate limits, crawl restrictions, and Crawl-delay values published by the site.
A successful crawl can make accurate product and service detail available to Alphalens's business discovery index. Pages that are blocked or absent from navigable links and sitemaps may be missing from that index. The operator does not promise a particular ranking, lead, quotation, or referral from any one crawl.
Alphalens describes this work as indexing for discovery, not collecting a corpus for generative model training. There is no documented training-data purpose for alphalens-bot. Site owners can therefore evaluate it as an AI search crawler without assuming that allowing it grants a separate AI training use.
Every request is documented with the exact user agent alphalens-bot. The operator's site links to public IP information, although this entry is not marked as IP verified and has no request-signing directory. Robots.txt remains the stated control: Alphalens says it checks the file before crawling and skips any URL that the file blocks.
