What it does
ClaudeBot is Anthropic's training crawler. Content is copied into datasets used to train foundation models. Blocking costs you nothing in traffic today; allowing is a donation of your content to model weights with no citation or referral in return.
Anthropic documents robots.txt + Crawl-delay support and no-CAPTCHA-circumvention. Blocking by IP can break the opt-out (impedes robots.txt reads) — prefer robots directives.
Full UA: not captured in our July 2026 roster review — match log hits on the token ClaudeBot (case-insensitive substring) and verify by IP before acting.
How to verify a hit is really ClaudeBot
Verification method (July 2026): Published IP list: https://claude.com/crawling/bots.json (fetched live 2026-08-22)
Strongest available check: the vendor publishes the crawler's egress IPs. Fetch the list, then confirm the hit's source IP falls inside it — a UA match from any other IP is a spoofer.
User-agent strings are freely spoofed, so a UA match alone proves nothing. Full step-by-step method: verify by IP.
Allow or block in robots.txt
Match the exact token ClaudeBot in robots.txt:
# Block ClaudeBot site-wide
User-agent: ClaudeBot
Disallow: /
# Explicitly allow ClaudeBot
User-agent: ClaudeBot
Allow: /
For the allow-search-block-training combined pattern (and the Allow-directive gotcha that silently kills carve-outs), see the robots.txt guide. robots.txt governs whether ClaudeBot may fetch; llms.txt is the separate, curated map AI assistants read once allowed in.
Vendor documentation: https://support.claude.com/en/articles/8896518-does-anthropic-crawl-data-from-the-web-and-how-can-site-owners-block-the-crawler
The Anthropic family
Anthropic operates 4 distinct tokens in this roster, and they do not block each other: a robots.txt line for ClaudeBot does nothing to its siblings — each needs its own User-agent: entry (the robots.txt guide has combined patterns).
| Token | Role | robots.txt |
|---|---|---|
| Claude-SearchBot | AI search | honors robots.txt |
| Claude-User | User action | claims robots compliance |
| anthropic-ai | Training | claims robots compliance |
Frequently asked questions
What is ClaudeBot?
ClaudeBot is Anthropic's training crawler. Content is copied into datasets used to train foundation models. Blocking costs you nothing in traffic today; allowing is a donation of your content to model weights with no citation or referral in return.
Does ClaudeBot respect robots.txt?
Yes — Anthropic documents robots.txt compliance for ClaudeBot (vendor docs reviewed July 2026).
How do I block ClaudeBot?
Add 'User-agent: ClaudeBot' followed by 'Disallow: /' to your robots.txt. Use the exact token — substring variants of other tokens will not match.
How do I verify ClaudeBot traffic by IP?
Fetch Anthropic's published IP list for ClaudeBot and confirm the hit's source IP falls inside it — a ClaudeBot user-agent from any other IP is a spoofer.
Related bots
Same vendor, then other training crawlers — the full directory profiles all 34.