Every AI bot · Apple
Applebot-Extended
Not a crawler. It's a name Apple reads in robots.txt to decide whether pages it collects for other reasons may be used to train AI. Blocking it doesn't stop Apple's other crawlers.
In robots.txt: User-agent: Applebot-Extended
The numbers
Apple's other bots
Sites often treat one company's bots differently, depending on what each is for.
| Bot | Purpose | Files naming it | Of those, block it |
|---|---|---|---|
| Applebot | Builds an AI search index | 37,767 | 11.3% |
Who does what
The highest-ranked sites on the top million in each group. Click one to see its whole robots.txt verdict.
Block it
- instagram.com #11
- fbcdn.net #13
- tiktok.com #54
- whatsapp.com #55
- msn.com #69
Block some paths
- facebook.com #3
- oracle.com #205
- ubuntu.com #233
- android.com #256
- amplitude.com #298
Name it and allow it
- wordpress.org #48
- ui.com #110
- kaspersky.com #140
- forter.com #270
- linktr.ee #272
Lists skip sites we don't showcase (adult content).
What this means for you: blocking Applebot-Extended is how you tell Apple not to use your pages for AI training, without leaving Apple's search. Decide each bot on what it's for, not who runs it.
Counts use one robots.txt per site on this month's Tranco top 1M (600,979 files we could read). Bot list: registry 2026-07-14.1, kept in sync with Dark Visitors and ai.robots.txt.