ClaudeBot: Anthropic's AI training crawler
Anthropic's crawler for public web content that may contribute to training its models.
- Free, no account.
- A few seconds.
- Nothing is stored.
ClaudeBot at a glance
Everything Anthropic documents about it, with links to the source.
- Operator
- Anthropic
- Type
- AI training
- Used for
- Training Anthropic's models
- Obeys robots.txt
- Yes. Anthropic documents that ClaudeBot follows robots.txt, so a rule for its token is honoured.
- robots.txt token
ClaudeBot- User-agent string
- Anthropic does not publish the full string. Match requests on the token ClaudeBot, and verify them against the IP list below.
- IP addresses
- claude.com/crawling/bots.json
- Documentation
- support.claude.com/en/articles/8896518
How to block ClaudeBot
Collects text that may be used to train AI models. Blocking it keeps your pages out of future training data. AI answers can still cite you through the search crawlers.
Block the whole site
Add this group to robots.txt at the root of your site.
User-agent: ClaudeBot Disallow: /
Block only some paths
List the paths; everything else stays open to ClaudeBot.
User-agent: ClaudeBot Disallow: /private/ Disallow: /drafts/
How to verify ClaudeBot
Anyone can send a crawler's user-agent string. Where the request comes from is what proves it.
- 1
Find requests whose user agent contains "ClaudeBot" in your server or CDN logs.
- 2
Check each source IP against Anthropic's published list (claude.com/crawling/bots.json). A request from outside it is not ClaudeBot, whatever it says.
- 3
Block what fails the check at the firewall. A robots.txt rule only reaches crawlers that choose to read it.
Other crawlers from Anthropic
Each has its own token, and blocking one leaves the others untouched.
Frequently asked questions
Anthropic's crawler for public web content that may contribute to training its models. Anthropic says ClaudeBot respects robots.txt, including the Crawl-delay extension, and does not try to bypass CAPTCHAs.
Yes. Anthropic documents that ClaudeBot follows robots.txt, so a rule for its token is honoured.
Add a group for its token to robots.txt at the root of your site: "User-agent: ClaudeBot" followed by "Disallow: /". To close only part of the site, list those paths instead of "/". To enforce the block rather than request it, deny the IP addresses Anthropic publishes at your firewall.
No. ClaudeBot is a training crawler: blocking it keeps your pages out of data that may train Anthropic's models. AI answers that cite pages come from search crawlers, which have their own tokens, so blocking ClaudeBot alone does not remove you from them.
Checked against Anthropic's documentation on 29 September 2026.
Next: see whether AI recommends you
Letting the right crawlers in is the first step. AskWatch shows what ChatGPT, Perplexity, Gemini and Google AI answer when buyers ask about your category, and who they name.
- Free, no credit card.
- Report in minutes, link sent to your email.