BotForYou Updated 9 October 2026 · Common Crawl CC-MAIN-2026-39

Every AI bot · Anthropic

Claude-SearchBot

Anthropic's crawler that builds an index for AI search answers.

In robots.txt: User-agent: Claude-SearchBot

The numbers

Named by1.8%of readable robots.txt files on the top 1M (10,987 of 600,979)
Of those, block it36.2%shut it out of the whole site (3,977 of 10,987)
Same sites, since CC-MAIN-2026-34+0.0 ptsshare of the same 394,690 sites that block it by name: 0.5% → 0.6%
Via Cloudflare's file2.9%of its blocks come from Cloudflare's ready-made robots.txt (115 of 3,977)
Block the whole site: 3,977Block some paths: 3,479Name it but allow it: 3,531
Every robots.txt file on the top 1M that names Claude-SearchBot, by what the rules that apply to it say.

Anthropic's other bots

Sites often treat one company's bots differently, depending on what each is for.

BotPurposeFiles naming itOf those, block it
ClaudeBotCollects pages to train AI70,49171.5%
anthropic-aiPurpose not documented28,01862.2%
claude-webFetches a page for a chatbot user21,07665.7%
Claude-UserFetches a page for a chatbot user9,95132.9%

Who does what

The highest-ranked sites on the top million in each group. Click one to see its whole robots.txt verdict.

Block it

  1. amazon.com #24
  2. tiktok.com #54
  3. msn.com #69
  4. nytimes.com #159
  5. cnn.com #200

Block some paths

  1. linkedin.com #17
  2. netflix.com #41
  3. yahoo.com #59
  4. snapchat.com #130
  5. flickr.com #189

Name it and allow it

  1. digicert.com #47
  2. godaddy.com #154
  3. forter.com #270
  4. hp.com #295
  5. lenovo.com #793

Lists skip sites we don't showcase (adult content).

What this means for you: blocking Claude-SearchBot can keep your pages out of Anthropic's AI search answers, including any links back to you. Decide each bot on what it's for, not who runs it.

Counts use one robots.txt per site on this month's Tranco top 1M (600,979 files we could read). Bot list: registry 2026-07-14.1, kept in sync with Dark Visitors and ai.robots.txt.