AI & Infrastructure
ClaudeBot and Claude-SearchBot Anthropic crawlers for AI model training and Claude search crawling

ClaudeBot / Claude-SearchBot (Anthropic Crawlers)

ClaudeBot and Claude-SearchBot are Anthropic web crawlers with different purposes: ClaudeBot collects web content that may contribute to AI model training, while Claude-SearchBot crawls and analyzes web content to improve Claude search results.
October 5, 2026
Cihan Geyik
Table of Content

ClaudeBot and Claude-SearchBot are web crawlers operated by Anthropic, but they serve different purposes. ClaudeBot collects public web content that could potentially contribute to the training of Anthropic's generative AI models, while Claude-SearchBot navigates and analyzes the web to improve the relevance and accuracy of search results for Claude users.

Anthropic also operates a third web agent called Claude-User, which can retrieve website content when a Claude user initiates a request.

Keeping these three identities separate is important for AI Crawler Accessibility because website owners can make different decisions about model training, search visibility, and user-directed retrieval.

ClaudeBot → Model Training
Claude-SearchBot → Search Quality & Search Indexing
Claude-User → User-Directed Web Retrieval

What Is ClaudeBot?

ClaudeBot is Anthropic's model-development crawler. Anthropic states that ClaudeBot helps improve the utility and safety of its generative AI models by collecting web content that could potentially contribute to their training.

This makes ClaudeBot fundamentally different from Claude-SearchBot. ClaudeBot's documented purpose concerns potential model-training data, rather than improving Claude's web search results.

ClaudeBot = model-training-related crawling. It should not be treated as Anthropic's search crawler.

What Is Claude-SearchBot?

Claude-SearchBot is Anthropic's web crawler for search. Anthropic states that it navigates the web to improve search result quality for users and analyzes online content to improve the relevance and accuracy of search responses.

Anthropic further explains that disabling Claude-SearchBot prevents its systems from indexing the site's content for search optimization. This may reduce the site's visibility and accuracy in user search results.

Claude-SearchBot = search-related crawling and indexing. For teams focused on visibility in Claude's web search experiences, this is the Anthropic crawler that requires particular attention.

ClaudeBot vs. Claude-SearchBot

ClaudeBot and Claude-SearchBot belong to the same Anthropic crawler ecosystem, but they should not be treated as interchangeable.

Anthropic Bot Documented Purpose Effect of Blocking
ClaudeBot Collect web content that could potentially contribute to generative AI model training Signals that future site materials should be excluded from Anthropic's AI model training datasets
Claude-SearchBot Analyze web content to improve search result relevance and accuracy Prevents Anthropic from indexing the content for search optimization and may reduce search visibility
Claude-User Retrieve web content in response to requests initiated by Claude users Prevents Anthropic from retrieving the site's content in response to those user queries
Training Policy ≠ Search Policy ≠ User-Directed Retrieval Policy

How to Allow ClaudeBot

A website that wants to explicitly permit ClaudeBot to crawl its public content can use a crawler-specific robots.txt group:

User-agent: ClaudeBot Allow: /

These instructions belong in the site's robots.txt file.

How to Block ClaudeBot

A website that does not want ClaudeBot crawling its content can use:

User-agent: ClaudeBot Disallow: /

Anthropic states that restricting ClaudeBot signals that future materials from the site should be excluded from its AI model training datasets.

This rule concerns ClaudeBot specifically. It does not automatically mean that Claude-SearchBot or Claude-User has been blocked.

How to Allow Claude-SearchBot

A site that wants to explicitly allow Anthropic's search crawler can use:

User-agent: Claude-SearchBot Allow: /

This communicates that Claude-SearchBot is permitted to crawl the public site under the applicable Robots Exclusion Protocol rules.

How to Block Claude-SearchBot

A website can block Claude-SearchBot with:

User-agent: Claude-SearchBot Disallow: /

Anthropic states that disabling Claude-SearchBot prevents its system from indexing the site's content for search optimization and may reduce the site's visibility and accuracy in user search results.

Blocking Claude-SearchBot is not the same as blocking ClaudeBot. The first affects Anthropic's search-related access; the second expresses a model-training preference.

Can You Block ClaudeBot but Allow Claude-SearchBot?

Yes. Because Anthropic identifies these bots separately, website owners can define separate robots.txt policies for them.

For example, a publisher that wants its content accessible for Claude search but does not want ClaudeBot crawling future content for potential training use could configure:

User-agent: Claude-SearchBot Allow: / User-agent: ClaudeBot Disallow: /

This distinction is particularly important for organizations that want to separate AI Search visibility from AI model-training permissions.

Claude-SearchBot Allowed → Search Crawling Allowed
ClaudeBot Blocked → Training-Related Crawling Restricted

What Is Claude-User?

Claude-User is Anthropic's user-directed web retrieval agent.

When someone asks Claude a question, Claude may access a website using the Claude-User agent to retrieve content in response to that request.

This differs from ClaudeBot's model-development crawling and Claude-SearchBot's search-quality crawling.

Claude User Asks a Question → Claude-User → Website Retrieval → Claude Response

Does Claude-User Respect Robots.txt?

Anthropic states that its bots honor industry-standard do not crawl directives in robots.txt. Its crawler documentation includes ClaudeBot, Claude-SearchBot, and Claude-User within this bot framework.

Anthropic also states that disabling Claude-User prevents its system from retrieving website content in response to a user query, which may reduce the site's visibility for user-directed web search.

This makes Claude-User an important additional consideration when evaluating Claude accessibility rather than focusing exclusively on Claude-SearchBot.

How to Control All Three Anthropic Bots Separately

A publisher can define a different policy for each Anthropic bot. For example:

User-agent: Claude-SearchBot Allow: / User-agent: Claude-User Allow: / User-agent: ClaudeBot Disallow: /

This configuration communicates three separate preferences:

  • allow Anthropic's search crawler;
  • allow user-directed retrieval through Claude-User; and
  • restrict ClaudeBot's model-training-related crawling.

The appropriate configuration depends on the website owner's own content, search, and AI training policies.

Do Anthropic Crawlers Respect Robots.txt?

Yes. Anthropic states that its bots respect "do not crawl" signals by honoring industry-standard robots.txt directives.

Anthropic also says its crawlers are designed to minimize disruption and that it supports the non-standard Crawl-delay extension to robots.txt.

User-agent: ClaudeBot Crawl-delay: 1

Crawl-delay can be useful when the goal is to reduce crawl frequency rather than block access entirely.

Anthropic Crawlers and Subdomains

Anthropic instructs website owners to apply their robots.txt preferences to every subdomain they want to opt out.

For example, policies for a main website may not automatically represent the intended policy for separate hosts such as:

  • www.example.com;
  • docs.example.com;
  • blog.example.com; and
  • app.example.com.

Teams auditing Crawler Access should therefore review all relevant hosts rather than checking only the primary domain.

Claude Crawlers and Crawl-delay

Anthropic explicitly supports the non-standard Crawl-delay robots.txt extension as a way to limit crawling activity.

This gives publishers an option between unrestricted crawling and a complete block when server load or crawl frequency is the primary concern.

Allow → Permit Crawling
Crawl-delay → Reduce Crawl Frequency
Disallow → Restrict Crawling

Do Anthropic Crawlers Bypass CAPTCHAs?

Anthropic states that its bots respect anti-circumvention technologies and that it will not attempt to bypass CAPTCHAs for the sites it crawls.

This also illustrates why AI Crawler Accessibility involves more than robots.txt.

A crawler may be allowed by robots.txt while technical infrastructure prevents successful retrieval.

Robots.txt Permission vs. Actual Claude Access

Allowing ClaudeBot or Claude-SearchBot in robots.txt does not guarantee that the crawler can successfully retrieve the site.

Other technical layers can include:

  • CDNs;
  • web application firewalls;
  • bot-management systems;
  • CAPTCHA challenges;
  • authentication;
  • IP restrictions;
  • geographic restrictions;
  • rate limits;
  • HTTP 403 responses;
  • HTTP 429 responses; and
  • server errors.
robots.txt → CDN / WAF → Bot Protection → Server → Content Retrieval

The difference between crawler permission and successful retrieval is a core part of Crawler Access.

How to Verify Anthropic Crawler Traffic

User-agent strings alone can be imitated, so identifying a bot name in server logs does not necessarily prove that the request originated from Anthropic.

Anthropic publishes a maintained list of source IP ranges for its crawlers. Anthropic states that a crawler request originating from an IP address on this list indicates that the crawler is coming from Anthropic.

The current list is available from Anthropic's crawler IP list.

Because infrastructure can change, teams should reference Anthropic's maintained list rather than permanently copying individual addresses into documentation.

Should You Block Anthropic Crawlers by IP Address?

Anthropic recommends using robots.txt to communicate crawler opt-out preferences.

Anthropic specifically warns that alternative approaches such as blocking crawler IP addresses may not work correctly or persistently guarantee an opt-out because IP blocking can prevent its bots from reading the site's robots.txt file.

For crawler preference management, use the documented robots.txt identities rather than treating IP blocking as a substitute for robots.txt.

Claude-SearchBot and AI Search Indexing

Claude-SearchBot is directly relevant to AI Search Indexing.

Anthropic states that disabling Claude-SearchBot prevents its system from indexing a site's content for search optimization.

However, indexing should not be treated as equivalent to appearing in every Claude answer.

Accessible → Crawled → Potentially Indexed → Potentially Retrieved → Potentially Used in a Response

Each stage is distinct, and successful crawler access alone does not guarantee downstream visibility.

Does Allowing Claude-SearchBot Guarantee Claude Visibility?

No. Allowing Claude-SearchBot removes one potential technical barrier to Anthropic's search crawler, but it does not guarantee that a particular page or brand will appear for a specific query.

Anthropic states that blocking Claude-SearchBot may reduce a site's visibility in user search results. That should not be reversed into a claim that allowing the crawler guarantees visibility.

Claude-SearchBot Allowed ≠ Guaranteed Claude Visibility

Does Allowing ClaudeBot Improve Claude Search Visibility?

ClaudeBot should not be treated as a Claude Search ranking mechanism.

Anthropic documents ClaudeBot for collecting content that could potentially contribute to model training, while Claude-SearchBot is the crawler specifically associated with search result quality.

ClaudeBot access and Claude Search visibility are different concepts.

Claude Crawlers and AI Visibility

For teams working on AI Visibility, the Anthropic crawler architecture creates an important distinction between technical eligibility and measured performance.

Claude-SearchBot accessibility can support search discovery, while Claude-User accessibility can affect user-directed retrieval. Neither automatically proves that a brand is appearing in relevant Claude answers.

Claude Accessibility → Search / Retrieval Opportunity → Claude Answers → Measured AI Visibility

Claude-SearchBot and AI Citations

Search crawler accessibility can be part of the technical foundation for source visibility, but it should not be used as a substitute for measuring actual AI Citations.

A page can be technically accessible without being selected as a source for a particular query.

Accessible ≠ Indexed ≠ Retrieved ≠ Selected ≠ Cited

This is why crawler configuration and citation monitoring should be evaluated as separate layers of an AI Search strategy.

How to Audit ClaudeBot and Claude-SearchBot

  1. Identify your policy. Decide separately whether you want to permit model-training crawling, search crawling, and user-directed retrieval.
  2. Review robots.txt. Check explicit rules for ClaudeBot, Claude-SearchBot, and Claude-User.
  3. Review every relevant subdomain. Do not assume one robots.txt file represents the policy for all hosts.
  4. Check important URL paths. Verify that public content intended for Claude search is not unintentionally blocked.
  5. Inspect infrastructure. Check CDN, WAF, bot protection, CAPTCHA, rate limits, and HTTP responses.
  6. Review crawl frequency. Consider Anthropic's Crawl-delay support when crawl load is the problem rather than access itself.
  7. Verify crawler traffic. Compare suspicious crawler traffic with Anthropic's current published crawler IP information.
  8. Measure downstream visibility. Do not assume successful crawling means successful Claude visibility.

Common Claude Crawler Mistakes

  • treating ClaudeBot and Claude-SearchBot as the same crawler;
  • blocking ClaudeBot and assuming Claude search is also blocked;
  • allowing ClaudeBot and assuming it improves Claude search rankings;
  • forgetting Claude-User when auditing user-directed retrieval;
  • allowing a bot in robots.txt while blocking it at the infrastructure layer;
  • checking only the main domain while ignoring subdomains;
  • using IP blocking as a replacement for documented robots.txt preferences;
  • assuming crawler access guarantees AI citations;
  • assuming indexing guarantees appearance in an answer; and
  • treating all Anthropic web traffic as model-training traffic.

Claude Crawlers and LLM SEO

Understanding Anthropic's crawler architecture is one technical component of LLM SEO, AEO, and Generative Engine Optimization (GEO).

The objective is not simply to allow every AI crawler. It is to define deliberate policies for different crawler purposes and ensure that search-related systems are not unintentionally prevented from accessing content intended for public discovery.

From Claude Crawler Access to Claude Visibility

Crawler access is an input. Actual visibility is the outcome that needs to be measured.

Ansvisor's Claude AI Visibility Tracker helps teams monitor how brands appear across relevant Claude prompts and answers.

Ansvisor's Citation Monitoring can then help measure which domains and URLs actually appear as sources instead of inferring citation performance from crawler configuration.

The broader Ansvisor AI Search Intelligence Platform connects AI visibility, prompts, citations, competitors, opportunities, and actions across the AI Search workflow.

Claude Crawler Access → Search & Retrieval → AI Visibility → Citations → Opportunities → Actions
ClaudeBot, Claude Bot, Claude-SearchBot, Claude SearchBot, Claude Search Bot, Anthropic Crawler, Anthropic Web Crawler, Anthropic Search Crawler, Claude AI Crawler, Claude Web Crawler

FAQ

Frequently asked questions.

What is ClaudeBot?

ClaudeBot is Anthropic's model-development crawler. Anthropic says it collects public web content that could potentially contribute to training its generative AI models. Restricting ClaudeBot signals that future materials from the site should be excluded from Anthropic's AI model training datasets.

What is Claude-SearchBot?

Claude-SearchBot is Anthropic's search crawler. It navigates and analyzes web content to improve the relevance and accuracy of search responses. Anthropic says disabling it prevents the system from indexing a site's content for search optimization and may reduce visibility in user search results.

What is the difference between ClaudeBot and Claude-SearchBot?

ClaudeBot is associated with collecting content that could potentially contribute to model training, while Claude-SearchBot is specifically associated with improving search result quality. Anthropic exposes them as separate crawler identities, allowing publishers to define different policies for each.

Can I block ClaudeBot but allow Claude-SearchBot?

Yes. Because Anthropic provides separate crawler identities, a website can disallow ClaudeBot while allowing Claude-SearchBot. This lets publishers distinguish model-training preferences from search crawler accessibility. Anthropic states that its bots honor robots.txt directives.

What is Claude-User and how is it different?

Claude-User supports user-directed web retrieval. When someone asks Claude a question, Claude may access a website using Claude-User. Anthropic says disabling Claude-User prevents its system from retrieving that site's content in response to a user query and may reduce visibility for user-directed web search.

Explore Ansvisor

Everything You Need to Improve Your AI Visibility

Track how your brand appears across AI platforms, understand what drives visibility, and turn insights into measurable actions.

From AI Visibility insights to action.

Explore the complete Ansvisor platform for AI Search intelligence, optimization, and growth.

Explore the Platform
Ansvisor is an open-source and cloud-ready AI Visibility Platform that helps brands measure, understand, and optimize their brand's AI visibility across ChatGPT, Claude, Gemini, Google AI Overviews, and other AI search platforms.

Win customers from all major AI platforms

Understand, measure, and optimize your AI visibility via Ansvisor.

✓ Add brand, domains and competitors
✓ Discover prompts and growth opportunities
✓ Track your AI visibility across major AI platforms
✓ Monitor citations, mentions, and competitors
✓ Measure AI traffic and customer discovery
✓ Receive AI recommendations based on AI insights
✓ Optimize authority, trust, and content quality
✓ Create content, automate analysis & action with AI agents

Help us grow the AI Visibility Grossary

New terms are added regularly.

Help us improve the page or suggest a new term →
About the Author
Cihan Geyik

Cihan Geyik

Co-founder at Ansvisor

Cihan Geyik is the co-founder of Ansvisor, an open-source AI Visibility platform for AI Search. With more than 15 years of experience in digital marketing and growth, he writes about AI visibility, AI search, AEO, GEO, citations, and answer engines. He focuses on helping brands understand and improve their presence across ChatGPT, Gemini, Perplexity, Google AI Overviews, and other AI-powered discovery platforms.

Summarize with ChatGPT
Summarize with Claude
Summarize with Google
Summarize with Perplexity
Summarize with Grok