Rankwise logoRankwise
Free audit
PricingCompareResources
Sign inGet started

Footer

Rankwise

The infrastructure layer for AI search: measure what AI says, publish what wins, and pinpoint the fixes that get you cited.

Built for the answer engines

Product

  • Pricing
  • Templates

Solutions

  • For agencies
  • For in-house SEO
  • For founders
  • Use Cases
  • Integrations
  • Compare
  • Alternatives

Resources

  • Resources
  • Articles
  • Guides
  • Case Studies
  • Learn
  • Glossary
  • Topics
  • API

Company

  • About
  • Contact
ImpressumAGBDatenschutz

© 2026 Rankwise. All rights reserved. Built by Founder Ventures.

Tracking ChatGPT · Perplexity · Claude · Google AI Overviews
  1. Rankwise
  2. /AI crawlers
  3. /ClaudeBot
Anthropic · Model training

ClaudeBot

Collects web content that could contribute to training Anthropic's models.

Source: privacy.claude.com
Check if your site allows ClaudeBot

Operator

Anthropic

Job

Model training

Collects pages for training future models. Slower to pay off, and separate from search and citations.

robots.txt

Follows robots.txt

Anthropic says its bots honour standard robots.txt directives, including the non-standard Crawl-delay. Blocking ClaudeBot excludes your future content from training data.

Source: privacy.claude.com
User agent

The token is all the vendor publishes

Anthropic documents the robots.txt token ClaudeBot but no full user-agent string. Match on the token.
robots.txt token
ClaudeBot
Source: privacy.claude.com
robots.txt

Allow or block ClaudeBot

Add one of these groups to the robots.txt at the root of your domain.

A crawler obeys only the most specific group that names it (RFC 9309), so a named group does not inherit the rules under User-agent: *. Repeat any paths you keep private, as the example does with /admin/.

Allow ClaudeBot
User-agent: ClaudeBot
Allow: /
Disallow: /admin/
Block ClaudeBot
User-agent: ClaudeBot
Disallow: /
What tryrankwise.com does

We name ClaudeBot and let it in

Read from the robots.txt we serve. We want AI engines to read and cite this site, so every AI crawler is allowed everywhere except the app, the API and the sign-in pages (/dashboard/, /api/, /login, /signup, /sign-in, /sign-up). The file also declares Content-Signal: ai-train=yes, search=yes, ai-input=yes.

tryrankwise.com/robots.txt
User-agent: ClaudeBot
Allow: /
Disallow: /dashboard/
Disallow: /api/
Disallow: /login
Disallow: /signup
Disallow: /sign-in
Disallow: /sign-up

See the whole file. The same rules repeat in every group, because a crawler that matches a named group ignores the wildcard one.

Worth knowing

More from Anthropic

  • Anthropic asks you to add the rule on every subdomain you want to opt out, and says blocking by IP address may not work reliably.

    Source: privacy.claude.com
Other Anthropic crawlers
Claude-SearchBotSearch index
Claude-UserUser-triggered fetch
AI engines it matters for
ClaudeHow it picks sources, and what Rankwise measures on it

Sources, checked 2026-09-24

  • privacy.claude.com/en/articles/8896518-does-anthropic-crawl-data-from-the-web-and-how-can-site-owners-block-the-crawler
Check your own file

Does your robots.txt let ClaudeBot in?

The free AI crawler checker reads your robots.txt and reports which of the major AI crawlers it allows or blocks. No signup.
Check AI crawler accessOr start free →