Domain explorer
Look up a single domain's robots.txt & llms.txt metrics, effective AI policy, inferred tech stack (labelled hypothesis), and crawl history. Read-only; one indexed lookup per query.
chatgpt.com
Tranco rank
68
popularity rank
TLD
com
chatgpt.com
robots.txt records
2
across crawls
llms.txt records
0
across crawls
Live files on chatgpt.com
Open the file as it exists on the host right now — may differ from what we measured. We link out rather than mirror (we store derived metrics, not raw bodies).
Effective AI / crawl policy (Google posture, §3.3)
| Crawl | Posture | Consec. 5xx | Reason | As of |
|---|---|---|---|---|
| CC-MAIN-2026-25 | blocked | 0 | 2xx: robots.txt disallows all crawling (Disallow: /) | 2026-06-18T13:26:54+00:00 |
robots.txt
| Crawl | Status | Bytes | Groups | Allow | Disallow | Sitemaps | Wildcard * | Disallow-all | AI tokens | AI verdicts |
|---|---|---|---|---|---|---|---|---|---|---|
| CC-MAIN-2026-25 | 200 | 3,970 | 14 | 144 | 20 | 5 | yes | yes | CCBot, img2dataset, Google-Extended, anthropic-ai, Claude-Web, Omgilibot, Omgili, FacebookBot, Bytespider, PerplexityBot | ccbot:block, img2dataset:block, google-extended:block, anthropic-ai:block, claude-web:block, omgilibot:block, omgili:block, facebookbot:block, bytespider:block, perplexitybot:block |
| CC-MAIN-2026-25 | 200 | 3,684 | 13 | 138 | 18 | 5 | yes | yes | CCBot, img2dataset, Google-Extended, anthropic-ai, Claude-Web, Omgilibot, Omgili, FacebookBot, Bytespider, PerplexityBot | — |
llms.txt
Collected by our own polite crawler, not Common Crawl — Common Crawl doesn't capture /llms.txt. See Methodology.
No llms.txt records for this domain.
Inferred tech stack hypothesis
No tech inference for this domain.