CrawlPact

Applebot-Extended

User-agent token
Applebot-Extended
Purpose
Training
Status
active
Last verified
2026-07-01

Applebot-Extended is documented by Apple as a control token for opting website content out of use in training Apple’s generative AI models, separate from the base Applebot crawler used for Siri and Spotlight indexing.

Site-owner controls

Disallowing Applebot-Extended opts this content out of training Apple Intelligence and other Apple generative AI models specifically. As with Google’s pairing of Googlebot/Google-Extended, it leaves standard Apple indexing for Siri and Spotlight Suggestions (Applebot) unaffected — the two tokens are evaluated independently, and Applebot-Extended does not itself perform a separate crawl; it layers a training-use restriction on top of Apple’s existing Applebot access. See /limitations for what a robots.txt rule can and cannot guarantee.

Example robots.txt configuration

To disallow Applebot-Extended specifically, without affecting any other crawler:

User-agent: Applebot-Extended
Disallow: /

If no dedicated User-agent: Applebot-Extended group exists in a domain's robots.txt, this crawler falls back to whatever the wildcard User-agent: * group says (RFC 9309) — see robots.txt syntax basics for how group selection works.

Official source: https://support.apple.com/en-us/119829

Verified against the source above as of 2026-07-01 — see how CrawlPact verifies crawler information.

See how this applies to your own site

Run a free audit to check your declared AI crawler policy against your own domain, or use the AI crawler checker to check this one crawler specifically.

Audit a domain

Analytics preferences

CrawlPact uses optional Google Analytics and Microsoft Clarity on public marketing pages to understand which content is useful and how visitors actually use it. Neither is used in the authenticated app or admin areas. You can accept or decline analytics without affecting the service.