AI training · Apple
Applebot-Extended
Not a crawler. A robots.txt token that tells Apple whether pages Applebot fetched may train Apple’s models. Disallowing it does not affect Siri or Spotlight.
- Operator
- Apple
- Purpose
- Model training. Crawlers that collect pages to train models. Blocking them does not affect search or AI answers.
- robots.txt token
Applebot-Extended- User agent
- None. This is a robots.txt token, not a crawler: the operator’s normal crawler fetches the page and this name only governs how it may be used.
- Documentation
- support.apple.com/en-us/119829
Block Applebot-Extended
Add a group naming it to /robots.txt. Nothing stops fetching; the operator’s crawler still reads the page for its normal purpose, but may not use it this way.
User-agent: Applebot-Extended
Disallow: /
Allow Applebot-Extended when everything else is blocked
A group naming the bot wins over User-agent: *, so an explicit allow lets it in while your catch-all rules stay in place.
User-agent: Applebot-Extended
Allow: /
User-agent: *
Disallow: /
Does Applebot-Extended reach your site?
The checker applies your robots.txt the way Applebot-Extended does, alongside 29 other bots.
Also run by Apple: Applebot.
Other ai training: GPTBot, ClaudeBot, Google-Extended, CCBot, GoogleOther, Meta-ExternalAgent, Bytespider.
Build a complete file with the robots.txt generator.