robots.txt & llms.txt across the Tranco top 1M
Data as of 2026-09-07 · robots.txt via Common Crawl CC-MAIN-2026-34 + our polite crawl · llms.txt via our polite crawl · Tranco list XN67N (2026-09-03T22:00:02.533886) · 1,000,000 panel domains 1,212,578 robots.txt fetches · 740,113 parsed · 56,498 llms.txt probes

Two public profiles · no score or winner

componentsearchengine.com and openfreemap.org

componentsearchengine.com #40,218 openfreemap.org #40,219

Measured takeaway

Botfy does not have comparable evidence for a strong policy takeaway; the available states are shown below.

Measured in Botfy's 2026-09-07 publication.

At a glance

A difference appears only when both sites have comparable evidence. Unavailable or unresolved observations remain incomplete.

Evidencecomponentsearchengine.com openfreemap.org
robots.txt Served and parsed Served and parsed
Google crawl posture Open Open
Unnamed crawler fallback Partly Restricted they inherit the wildcard (*) group's rules Allowed the wildcard (*) group is effectively open
Base /llms.txt Not probed Not found when probed

Treatment by purpose

Only resolved blocked, allowed, or mixed states count as a policy agreement or difference. Unnamed agents still inherit each site's wildcard rules.

Declared purposecomponentsearchengine.com openfreemap.org
Model training Not explicitly addressed Not explicitly addressed
AI search Not explicitly addressed Not explicitly addressed
Live assistant fetch Not explicitly addressed Not explicitly addressed
Opt-out directives Not explicitly addressed Not explicitly addressed
Ambiguous / undocumented Not explicitly addressed Not explicitly addressed

Named agents

Measured differences appear first. “Not explicitly addressed” is not the same as allowed; the site's wildcard fallback governs that crawler.

Agentcomponentsearchengine.com openfreemap.org
gptbot OpenAI Not explicitly addressed Not explicitly addressed
claudebot Anthropic Not explicitly addressed Not explicitly addressed
oai-searchbot OpenAI Not explicitly addressed Not explicitly addressed
perplexitybot Perplexity Not explicitly addressed Not explicitly addressed
chatgpt-user OpenAI Not explicitly addressed Not explicitly addressed
claude-user Anthropic Not explicitly addressed Not explicitly addressed

How each site compares with its own rank peers

The sites may belong to different rank bands. Each percentage keeps its own population and denominator; Botfy does not subtract rates from unlike cohorts.

Traitcomponentsearchengine.com openfreemap.org
Wildcard group Yes 92.1% Ranks 10,001–100,000 · n=54,147 Yes 92.1% Ranks 10,001–100,000 · n=54,147
Effectively open file No 22.6% Ranks 10,001–100,000 · n=54,147 Yes 22.6% Ranks 10,001–100,000 · n=54,147
Sitemap declared Yes 61.7% Ranks 10,001–100,000 · n=54,147 Yes 61.7% Ranks 10,001–100,000 · n=54,147
robots.txt served Yes 59.7% Ranks 10,001–100,000 · n=91,507 Yes 59.7% Ranks 10,001–100,000 · n=91,507
robots.txt returned 404 No 7.9% Ranks 10,001–100,000 · n=91,507 No 7.9% Ranks 10,001–100,000 · n=91,507
Over the RFC parse limit No 0.2% Ranks 10,001–100,000 · n=54,147 No 0.2% Ranks 10,001–100,000 · n=54,147
Blanket disallow No 4.0% Ranks 10,001–100,000 · n=54,147 No 4.0% Ranks 10,001–100,000 · n=54,147
How this comparison is measured

Both sites use Botfy's latest published observations. A missing fetch, unresolved agent verdict, and measured absence are different states. Peer percentages use each site's exclusive Tranco rank band and print their own denominator. Populations, sources, and limits →

Keep comparing

Each suggestion replaces one side using an editorial collection or current rank proximity. Neither means the sites have similar policies.