Bot directory · AI training

AI training · Apple

Applebot-Extended

Not a crawler. A robots.txt token that tells Apple whether pages Applebot fetched may train Apple’s models. Disallowing it does not affect Siri or Spotlight.

Operator
Apple
Purpose
Model training. Crawlers that collect pages to train models. Blocking them does not affect search or AI answers.
robots.txt token
Applebot-Extended
User agent
None. This is a robots.txt token, not a crawler: the operator’s normal crawler fetches the page and this name only governs how it may be used.
Documentation
support.apple.com/en-us/119829

Block Applebot-Extended

Add a group naming it to /robots.txt. Nothing stops fetching; the operator’s crawler still reads the page for its normal purpose, but may not use it this way.

User-agent: Applebot-Extended
Disallow: /

Allow Applebot-Extended when everything else is blocked

A group naming the bot wins over User-agent: *, so an explicit allow lets it in while your catch-all rules stay in place.

User-agent: Applebot-Extended
Allow: /

User-agent: *
Disallow: /

Does Applebot-Extended reach your site?

The checker applies your robots.txt the way Applebot-Extended does, alongside 29 other bots.

Also run by Apple: Applebot.

Other ai training: GPTBot, ClaudeBot, Google-Extended, CCBot, GoogleOther, Meta-ExternalAgent, Bytespider.

Build a complete file with the robots.txt generator.