All free tools

Check technical SEO · Free · No signup

Robots.txt Checker

Check crawler directives and common accidental blocking patterns.

Public information only. Do not enter passwords, private URLs, or confidential data.

What this robots.txt checker checks and how to act on it

This robots.txt checker fetches the live robots.txt from the host of the URL you enter, parses every User-agent group, and tests that exact URL against Googlebot and against a generic crawler using longest-match rules. It reports the winning rule and its line number, so an unexpected block can be traced to the line that caused it. It also flags a site-wide Disallow for all crawlers, blocked CSS, JavaScript, or image paths, Crawl-delay values, lines crawlers cannot parse, a missing Sitemap directive, and whether AI crawlers have groups of their own.

Three failures matter most. A server error on robots.txt makes Google treat the whole site as disallowed until the file recovers. A Disallow rule left over from staging can block the entire site or its CSS and JavaScript paths, starving the renderer. And AI crawler groups for GPTBot, ClaudeBot, PerplexityBot, and Google-Extended can be blocked by a copied template or a CDN default, which ends any chance of citation in AI answers.

Run the checker after every deploy and whenever a CDN or CMS setting changes, one URL per run: the homepage, a key landing page, a representative asset path, and one URL you expect to be blocked. For a per-bot verdict across GPTBot, ClaudeBot, PerplexityBot, and other AI crawlers, use the AI crawler access checker. robots.txt controls crawling, not indexing: a page blocked here can still appear in results if it is linked elsewhere, so use a noindex tag to keep a page out of the index.

What this checker cannot tell you: whether a CDN or firewall returns 403 to a crawler before robots.txt is consulted, whether an allowed page actually contains its content in the static HTML, or whether a blocked page is still indexed from links. Pair it with the AI crawler access checker, a fetch with JavaScript disabled, and Search Console's Pages report.

How to use the robots.txt checker

  1. Enter the full URL of a page you want to test, such as your homepage or a key landing page.
  2. The checker fetches robots.txt from that host and tests the URL against Googlebot and a generic crawler.
  3. Read the winning rule: its line number shows which Allow or Disallow decided the result.
  4. Fix any unexpected block in robots.txt, deploy, and run the check again.

What it does not check

  • Firewall or CDN rules that return 403 to a crawler before robots.txt is read.
  • Whether the page is indexed; robots.txt controls crawling, not indexing.
  • A verdict for each AI crawler; the AI crawler access checker covers that.

Read the guide

Step-by-step explanations of what this tool checks and how to fix what it finds.

Related free tools

Choose how SerionFlow uses cookies.

Essential cookies keep the site secure and working. Product analytics helps us improve reliability. Optional cookies help us remember preferences and improve campaigns. You can change optional cookies anytime in Cookie Settings.