What is Amazon Q?
Amazon Q Business is a generative AI assistant that an organization can configure around its own information. The Web Crawler connector is one of the data sources available to that assistant. It gathers pages selected by the organization and adds them to the application's searchable content.
The connector accepts public sites and internal company sites over HTTPS. A customer can start from URLs or sitemaps, and AWS documents authentication and proxy options for content that is not openly reachable. AWS also requires customers to crawl only their own pages or pages they have permission to index.
An Amazon Q request therefore points to a particular customer's configuration. It is not evidence that Amazon has chosen the page for a public web index. The content may later be used to answer questions from people who have access to that customer's Q Business application.
This has a narrow visibility effect. Blocking the crawler can keep a page out of the relevant organization's assistant, but it does not lower the page in Google, Bing, or public AI search. AWS presents the crawl as retrieval-data ingestion and does not document it as a foundation-model training crawl.
AWS's robots documentation currently names amazon-QBusiness for its Web Crawler. The stored bot record does not define a stable user-agent token, and there is no matching Cloudflare directory entry or Web Bot Auth key URL. Confirm the live request against the current AWS connector documentation before making an identity-based firewall exception.
AWS now notes that Amazon Q Business is no longer open to new customers. Existing applications and configured web data sources can still account for this traffic, so the product status alone is not a reason to label a request fake. The application owner should be able to identify the data source and its authorized scope.
