Promptwatch Logo

Applebot-Extended

Applebot-Extended is a robots. txt control token, not a crawler that makes its own page requests. Apple calls it a secondary user agent.
AppleApplebot-Extended
AI Training

What is Applebot-Extended?

Applebot-Extended is a robots.txt control token, not a crawler that makes its own page requests. Apple calls it a secondary user agent. Its only job is to tell Apple how content already fetched by Applebot may be used.

Applebot performs the actual crawl. When Apple processes a site's rules, an Applebot-Extended directive can prevent the resulting data from being used to train Apple's general-purpose foundation models. Apple's Applebot documentation is the source for this separation.

The training scope includes generative features across Apple products, including Apple Intelligence, Services, and Developer Tools. A disallow rule is therefore a specific training-data opt-out. It is not a request to remove the page from Apple's search systems.

Apple says pages covered by an Applebot-Extended disallow can still appear in search results. The rule does not block Applebot itself and does not remove eligibility for search experiences in Spotlight, Siri, or Safari. Publishers can make the training choice without giving up those search surfaces.

Because Applebot-Extended never crawls pages, it should not appear as a separate visitor in access logs. A log entry will normally identify Applebot, while the Extended decision comes from the site's robots file. There is no independent Applebot-Extended request to authenticate or rate-limit.

Applebot-Extended is recorded as following robots.txt and has a stable token. Rules can cover the whole site or selected public paths. Allowing the token permits the documented training use, but it does not guarantee that any model will learn a particular fact or mention the site in generated output.

Relevant for AI searchApple Intelligence

Is Applebot-Extended relevant for AI search?

Yes. Applebot-Extended feeds Apple Intelligence, so the pages it can reach shape what those AI products say about you.

Applebot-Extended gathers public web content that can end up in the training data for large language models. Once your pages are in that set, they influence how the operator's models talk about you for that model generation. Allowing it lets your own writing carry weight in those answers; blocking it means the models learn about you from third parties instead.

How to handle Applebot-Extended

Decide this policy as a training-data question. To opt the entire site out, address the stable token in robots.txt:

User-agent: Applebot-Extended
Disallow: /

Use a path instead of / when only one public section should be excluded. Keep any Applebot crawl rules separate; blocking Applebot-Extended alone does not stop Apple search crawling.

Check the live robots file after deployment. Do not look for Applebot-Extended page requests as proof, since Apple states that this secondary user agent does not crawl webpages.

Examples

  • A news publisher disallows Applebot-Extended across its public archive while continuing to allow Applebot for Apple search results.
  • A software vendor excludes `/documentation/enterprise/` from Apple's foundation-model training without changing the rest of the site's policy.
  • A webmaster sees only Applebot in access logs and checks robots.txt, rather than waiting for a separate Applebot-Extended visit that will never occur.

Frequently asked questions about Applebot-Extended

Learn about AI visibility monitoring and how Promptwatch helps your brand succeed in AI search.

Apple defines and operates Applebot-Extended as a secondary user-agent control for data fetched by Applebot.

Be the brand AI recommends

Monitor your brand's visibility across ChatGPT, Claude, Perplexity, and Gemini. Get actionable insights and create content that gets cited by AI search engines.

Promptwatch Dashboard