Bot directory · AI training

AI training · Anthropic

ClaudeBot

Collects pages that may be used to train Anthropic’s models. Blocking it does not affect Claude’s search or user requests.

Operator
Anthropic
Purpose
Model training. Crawlers that collect pages to train models. Blocking them does not affect search or AI answers.
robots.txt token
ClaudeBot
User agent
Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; ClaudeBot/1.0; +claudebot@anthropic.com)
Documentation
support.claude.com/en/articles/8896518

Block ClaudeBot

Add a group naming it to /robots.txt. Compliant crawlers stop before requesting any page. A firewall rule on the user agent is the only way to enforce it against crawlers that ignore robots.txt.

User-agent: ClaudeBot
Disallow: /

Allow ClaudeBot when everything else is blocked

A group naming the bot wins over User-agent: *, so an explicit allow lets it in while your catch-all rules stay in place.

User-agent: ClaudeBot
Allow: /

User-agent: *
Disallow: /

Does ClaudeBot reach your site?

The checker applies your robots.txt the way ClaudeBot does, then requests your page with the user agent above, alongside 29 other bots.

Also run by Anthropic: Claude-SearchBot, Claude-User.

Other ai training: GPTBot, Google-Extended, CCBot, Applebot-Extended, GoogleOther, Meta-ExternalAgent, Bytespider.

Build a complete file with the robots.txt generator.