What is Amazon Kendra?
Amazon Kendra is AWS's managed search and information-retrieval service. Organizations use it to build search over documents they choose, with natural-language queries and answers drawn from the indexed material. It is an enterprise search product, not a public assistant that independently surveys the web.
Kendra's Web Crawler is one way a customer can feed pages into that index. The customer supplies seed URLs or sitemaps and chooses the crawl scope. AWS supports public and internal HTTPS sites, along with authenticated access and web proxies for some connector configurations.
AWS requires customers to crawl only pages they own or are authorized to index. That matters when this user agent appears on an unrelated public site: the request reflects one customer's Kendra data source, not an Amazon-wide collection program. The site may be destined for an internal knowledge search used by that customer's staff or users.
A successful crawl can make a page searchable inside the configured Kendra index. It does not place the page in a general Amazon search engine, ChatGPT, or another public AI answer service. AWS documentation describes the workflow as document indexing and retrieval, not collection for training a generative foundation model.
Cloudflare identifies Amazon as the operator and links to the Amazon Kendra documentation. Its directory records user agents such as amazon-kendra-customer-id-[id] and amazon-kendra-web-crawler-*, with the shared pattern amazon-kendra-. The customer-specific portion can vary between requests.
There is no single stable user-agent token stored for this page, and no Web Bot Auth key directory is listed. Treat the user agent as a classification clue, not proof that Amazon or a particular Kendra customer authorized the crawl. The requested host, path selection, timing, and responses are useful when investigating who configured the data source.
