Crawl & Access

robots.txt Checker

Parse allow and disallow rules per crawler and highlight conflicts.

Free · instant

Run a deterministic technical check

The result comes from HTTP requests, parsers and maintained rules — no AI guessing, so it is reproducible.

The check runs against the site root, e.g. yourbrand.com
01

What does this tool check?

We fetch /robots.txt, parse every User-agent group with the standard longest-match rule, and show which groups block the site root, which sitemaps are declared, and which lines no parser understands.

02

How should I read the result?

A missing robots.txt means "allow everything" — that is a valid state, not an error. robots.txt controls crawl permission; it does not guarantee indexing or citation.

03

What should I do next?

If you intend to allow AI crawlers, make sure no User-agent group accidentally matches them with a Disallow: /.

Sources: RFC 9309 Robots Exclusion Protocol

InsightWonder

Access rules change without warning

robots.txt edits and CDN bot rules usually ship without anyone telling you. InsightWonder re-checks on a schedule and alerts you when a crawler starts getting blocked.

Run a full analysis
  • Scheduled re-checks with change alerts
  • Visibility and citations across five AI engines
  • Fixes ranked by impact, with content generated
robots.txt Checker for AI Crawlers — Free Rule Tester