Anthropic runs three web crawlers, and only one of them is about training. ClaudeBot collects content that may be used to train models. Claude-SearchBot indexes pages to improve Claude's search results. Claude-User fetches a page live when someone asks Claude a question that needs it. If you want Claude to cite you, allow the second and third. Blocking the first keeps you out of future training data, and Anthropic's crawler documentation (last updated 7 April 2026) describes it separately from search.
What each Anthropic crawler does, in Anthropic's words
ClaudeBot. It "helps enhance the utility and safety of our generative AI models by collecting web content that could potentially contribute to their training." Restricting it "signals that the site's future materials should be excluded from our AI model training datasets."
Claude-SearchBot. It "navigates the web to improve search result quality for users." Blocking it "prevents our system from indexing your content for search optimization, which may reduce your site's visibility and accuracy in user search results."
Claude-User. "When individuals ask questions to Claude, it may access websites using a Claude-User agent." Blocking it "prevents our system from retrieving your content in response to a user query, which may reduce your site's visibility for user-directed web search."
Two details in that wording matter. First, Anthropic says "may reduce", not "removes": Claude-SearchBot is not described as the only route into Claude's answers (see the Brave section below). Second, Claude-User is not just for pasted links. It fetches pages to answer questions, so blocking it can cut you out of answers even when your page is indexed.
The same document says all three respect robots.txt and honor the non-standard Crawl-delay directive. Anthropic asks site owners not to block by IP address, because that also stops its crawlers from reading robots.txt. It publishes its crawler IP ranges at claude.com/crawling/bots.json for allowlisting.
The robots.txt that keeps you citable
For a site that wants out of training but into Claude's answers:
User-agent: ClaudeBot
Disallow: /
User-agent: Claude-SearchBot
Allow: /
User-agent: Claude-User
Allow: /
Claude-SearchBot and Claude-User were documented in 2025, so older "block AI crawlers" templates never named them. Those templates still catch them through a broad User-agent: * disallow, or through rules aimed at "Anthropic" in general. Read your file for both.
Where Claude's search results come from
Partly from Brave Search. In March 2025 Anthropic added Brave Search to its list of subprocessors, and developers found a BraveSearchParams parameter in Claude's web search, plus at least one search where Claude and Brave returned identical citations (TechCrunch, 21 March 2025). Anthropic has not published how Brave results and its own Claude-SearchBot index are combined, so treat them as two inputs.
The Brave link has a consequence most Claude guides miss. Brave's crawler does not announce its own user agent, and Brave says that "if a domain or page is not crawlable by Googlebot, then Brave Search's bot will not crawl it either." A rule that blocks Googlebot from a section of your site therefore keeps that section out of Brave, and plausibly out of Claude's search results. Brave also says robots.txt is not how you keep a page out of its index; it uses the noindex directive for that.
A quick self-check: search site:yourdomain.com on search.brave.com and confirm your important pages appear.
How big is Claude as a traffic source?
Small, and growing fast. In SE Ranking's panel of 101,574 sites (published June 2026), Claude referral traffic grew 386% between January and April 2026, yet it still accounted for 1.40% of AI-referred traffic, against 78.23% for ChatGPT. Similarweb reported Claude's US monthly users up 349% between June 2025 and May 2026 (PPC Land).
Web search itself is not new: Anthropic launched it on 20 March 2025 for paid US users and made it available on all plans worldwide on 27 May 2025 (Claude blog). Developers building on the Claude API have to switch web search on as a paid tool, so not every product built on Claude searches the web. When one does show your page, the visit carries that product's referrer, not claude.ai's. Your analytics will undercount Claude for that reason, by an amount nobody can measure.
What is not known about Claude citations
Anthropic publishes no citation statistics. None of the large public studies of where on a page engines quote from, or how fresh the cited pages are, included Claude. One that did include it, SparkToro and Gumshoe's January 2026 test of about 3,000 runs across ChatGPT, Claude and Google's AI, found the same list of recommendations came back less than once in 100 runs (Search Engine Land). So a single Claude check tells you little. Measure a rate, as described in how to track AI citations.
Anything more specific you read, such as how often Claude searches or how strictly it filters sources, is practitioner impression. The page-level work that applies to every engine (direct answers, self-contained sections, attributed facts) is covered in what is AEO.
Your Claude checklist
- robots.txt:
Claude-SearchBotandClaude-Userallowed, including under anyUser-agent: *rule.ClaudeBotas you choose. - Googlebot not blocked from anything you want Claude to cite, because Brave follows Googlebot's rules.
- No IP-level blocks on Anthropic ranges; CDN AI-bot settings checked.
- Server logs show
Claude-SearchBotandClaude-Userrequests with 200 responses. - Your key pages show up for
site:searches on search.brave.com. - Ten real questions run in claude.ai with web search on, repeated over a few days, with the cited domains recorded.
Can I slow Anthropic's crawlers down instead of blocking them?
Does blocking ClaudeBot remove content Anthropic has already collected?
Sources
- Anthropic — Does Anthropic crawl data from the web, and how can site owners block the crawler? (updated 7 Apr 2026)
- Anthropic — Claude crawler IP ranges (bots.json)
- Claude blog — Web search launch and global availability
- TechCrunch — Anthropic appears to be using Brave to power web searches for its Claude chatbot, 21 Mar 2025
- Brave — Brave Search crawler
- Cloudflare — Content Independence Day: AI crawler options, 1 Jul 2026
- SE Ranking — Claude traffic research, 3 Jun 2026
- PPC Land — Similarweb data on ChatGPT, Gemini and Claude web share, Jul 2026
- Search Engine Land — SparkToro/Gumshoe: AI recommendation lists rarely repeat, Jan 2026
Updated October 1, 2026: rebuilt around Anthropic's own descriptions of ClaudeBot, Claude-SearchBot and Claude-User; corrected the claim that blocking Claude-User only affects pasted links; added Claude's Brave Search dependency and the Googlebot rule that comes with it.