robots.txt & llms.txt across the Tranco top 1M

Domain explorer

Look up a single domain's robots.txt & llms.txt metrics, effective AI policy, inferred tech stack (labelled hypothesis), and crawl history. Read-only; one indexed lookup per query.

snapchat.com

Tranco rank
119
popularity rank
TLD
com
snapchat.com
robots.txt records
2
across crawls
llms.txt records
1
across crawls

Live files on snapchat.com

Open the file as it exists on the host right now — may differ from what we measured. We link out rather than mirror (we store derived metrics, not raw bodies).

Effective AI / crawl policy (Google posture, §3.3)

CrawlPostureConsec. 5xxReasonAs of
CC-MAIN-2026-25 open 0 2xx: robots.txt present, no full disallow in effect 2026-06-18T14:15:02+00:00

robots.txt

CrawlStatusBytesGroupsAllowDisallow SitemapsWildcard *Disallow-allAI tokensAI verdicts
CC-MAIN-2026-25 200 4,490 51 8 61 19 yes no Meta-ExternalAgent, Meta-ExternalFetcher, OAI-SearchBot, ChatGPT-User, PerplexityBot, Perplexity-User, Claude-SearchBot, Claude-User, DuckAssistBot, AI2Bot, Ai2Bot-Dolma, Amazonbot, anthropic-ai, Applebot-Extended, Bytespider, CCBot, Claude-Web, ClaudeBot, cohere-ai, cohere-training-data-crawler, Diffbot, FacebookBot, Google-Extended, GoogleOther, GPTBot, img2dataset, Kangaroo Bot, omgili, omgilibot, PetalBot, Scrapy, Timpibot, Webzio-Extended, YouBot meta-externalagent:block, meta-externalfetcher:block, oai-searchbot:partial, chatgpt-user:partial, perplexitybot:partial, perplexity-user:partial, claude-searchbot:partial, claude-user:partial, duckassistbot:partial, ai2bot:block, ai2bot-dolma:block, amazonbot:block, anthropic-ai:block, applebot-extended:block, bytespider:block, ccbot:block, claude-web:block, claudebot:block, cohere-ai:block, cohere-training-data-crawler:block, diffbot:block, facebookbot:block, google-extended:block, googleother:block, gptbot:block, img2dataset:block, kangaroo bot:block, omgili:block, omgilibot:block, petalbot:block, scrapy:block, timpibot:block, webzio-extended:block, youbot:block
CC-MAIN-2026-25 200 4,308 49 1 55 21 yes no Meta-ExternalAgent, Meta-ExternalFetcher, AI2Bot, Ai2Bot-Dolma, Amazonbot, anthropic-ai, Applebot-Extended, Bytespider, CCBot, ChatGPT-User, Claude-Web, ClaudeBot, cohere-ai, cohere-training-data-crawler, Diffbot, DuckAssistBot, FacebookBot, Google-Extended, GoogleOther, GPTBot, img2dataset, Kangaroo Bot, OAI-SearchBot, omgili, omgilibot, PerplexityBot, PetalBot, Scrapy, Timpibot, Webzio-Extended, YouBot

llms.txt

Collected by our own polite crawler, not Common Crawl — Common Crawl doesn't capture /llms.txt. See Methodology.

CrawlStatusVariantConformantSections LinksGenerator
CC-MAIN-2026-25 404

llms.txt data reflects adoption + conformance only, never consumption — publishing a file does not mean any AI reads it.

Inferred tech stack hypothesis

No tech inference for this domain.