SEO

Robots.txt Tester

Test any URL against a robots.txt file for any user agent, showing which rule matched and why, with the longest-match precedence rule applied exactly as crawlers apply it.

Last reviewed by the Radiatus Cloud team

Results appear here.

Need this done properly for your business?

Radiatus delivers secure cloud, DevOps & compliance engineering.

Book a free consult

The matching rule is not first match wins

The rule that applies is the longest matching path, not the first one in the file, and when an allow and a disallow of the same length both match, allow wins. This surprises people repeatedly: a file with Disallow slash followed by Allow slash blog does permit the blog, because the allow rule is longer for those URLs. Reading a robots file top to bottom and stopping at the first match gives the wrong answer in exactly the cases where the file is doing something interesting.

Only one group applies

A crawler uses the single most specific group matching its user agent and ignores every other group entirely, including the wildcard group. If Googlebot has its own section, the rules under the wildcard do not apply to it at all, even for paths its own section never mentions. Adding a specific group for one crawler therefore silently removes all the wildcard rules from that crawler, which is one of the most common ways a robots file stops doing what its author intended.

Disallow does not mean deindex

A blocked URL can still appear in search results if other pages link to it, because the crawler knows the URL exists without being permitted to fetch it. The listing shows without a description, since the content was never read. To keep a page out of the index it must be crawlable and carry a noindex directive, which means blocking it in robots actually prevents removal by making the noindex tag unreadable.

Related tools

  • Meta Tag Analyzer — Check a page's meta tags for the problems that actually affect indexing and click-through.
  • Redirect Chain Checker — Trace the full path of redirects (301/302) to find loops or lost link juice.
  • Robots.txt Validator — Validate robots.txt syntax and check whether a specific URL is allowed or blocked for a given crawler.
  • Sitemap Generator (Lite) — Generate a valid XML sitemap with lastmod dates and correct structure, and learn which URLs belong in it and which do not.

Frequently Asked Questions

Which rule wins when several match?

The longest matching path, regardless of order in the file. If an allow and a disallow of equal length both match, the allow wins. Reading top to bottom and stopping at the first match gives the wrong answer for any non-trivial file.

Does a specific user agent group inherit the wildcard rules?

No. A crawler uses only the most specific group matching its name and ignores every other group entirely. Adding a section for one crawler silently removes all the wildcard rules from it, which is a very common cause of unintended behaviour.

Does blocking a URL remove it from search results?

No. A blocked URL can still be listed if other pages link to it, shown without a description because the content was never fetched. Removal requires a noindex directive on a page the crawler is allowed to read, so blocking it actually prevents removal.

Are wildcards supported?

The asterisk for any sequence of characters and the dollar sign for end of URL are supported by the major crawlers, though they are extensions rather than part of the original standard. This tester implements them the way the major crawlers do.

Is robots.txt case sensitive?

The paths are, since URLs are. The directive names and user agent names are not. A rule disallowing /Admin does not block /admin, which is a frequent and easily missed mistake.

Privacy & Security

Everything runs in your browser; nothing is uploaded.

Data: None
Client-side-Side
Active
v1.0

How to Use

Paste a robots.txt and a list of URLs to see which are allowed and which rule decided it.

Disclaimer: This tool is provided "as is" without warranty of any kind. Results are for educational and utility purposes.