StayIndexed
Free tool

Robots.txt check (Robots.txt Checker)

Enter a site address: we will download robots.txt, parse it by Google's rules, find errors and check whether the page you need is blocked for a bot. No sign-up.

Try:

The check runs from our server: we download robots.txt the way a search bot does. The URL field is optional: fill it in to find out whether the selected bot is allowed to crawl a specific page. We do not store your queries together with your IP address.

Your first 30 tokens are on us

The file is fine. But are the pages themselves in Google’s index?

robots.txt only allows crawling; it does not tell you whether pages made it into search. StayIndexed checks that for every URL.

  • Google indexing check for any URL, with the technical reason if the page is not in the index
  • Bulk checks of hundreds and thousands of URLs at once
  • Regular monitoring and an alert the same day a page drops out of the index
30
tokens at sign-up — that is 10 free checks
Try it for free Plans and pricing

No card · sign in with Google · 1 check = 3 tokens, from $0.0016

What Robots.txt Checker checks

How robots.txt works

The file lives in the site root (https://example.com/robots.txt) and consists of groups: a User-agent line names the bot, and below it come Allow and Disallow rules with paths. A bot picks the group that fits it best, and among the rules the one with the longer path wins; if the length is equal, Allow wins. The characters * (any characters) and $ (end of the address) are supported.

Common robots.txt errors

Disallow: / after the site launchA site blocked during development is left blocked after the release, and search traffic disappears. This is the most expensive mistake.
Blocked CSS and JSGoogle renders pages like a browser. If styles and scripts are blocked, it sees a "broken" page and rates it lower.
Noindex in robots.txtGoogle has not supported the Noindex directive since 2019. To remove a page from search, use meta robots or X-Robots-Tag.
A relative Sitemap addressAfter Sitemap: a full address with https:// is needed. An entry like /sitemap.xml may be ignored by a bot.

Does Disallow hide a page from Google

No. Disallow forbids a bot to crawl a page, but does not forbid showing it in search: if there are links to it, the address can remain in the results without a description. To remove a page from the index, it must be left open to crawling and have noindex added. Whether a page has got into Google is shown by the page indexing check.

Limits of the check

We download the file from a public address with one request, so sites that block third-party bots (Cloudflare, etc.) may return an error to us even though Googlebot sees the file normally. Each subdomain and protocol has its own robots.txt, so check the host you need separately. The sitemap file that the Sitemap: line points to can be checked by Sitemap Checker.

How a page opens after redirects is shown by Redirect Checker.

What else StayIndexed can do

This page is a free tool from the StayIndexed service. In your account you can save a list of your sites, refresh the check in one click and see the history of robots.txt changes (also free), while the main job of the service is to track how your pages are doing in Google.

How much it costs

The check is free, both on this page and in your account. Tokens are used to pay for the other services: indexing, speed and backlinks cost 3 tokens per check. 30 tokens are credited at sign-up, no card required.

Up to 10000 tokens$0.00102
10001–50000$0.00068
Over 50000$0.00054
Buy tokens

Frequently asked questions

What is robots.txt?

It is a text file in the site root that tells search bots which sections may be crawled and which may not, and where the sitemap is. It is a recommendation for bots, not protection against access.

Where should robots.txt be located?

Only in the root of the host: https://example.com/robots.txt. Each subdomain and protocol (http and https) has its own file, and a file in a subfolder is not read by bots.

Does Disallow hide a page from Google?

No. Disallow forbids crawling, but the address can still get into search without a description. To remove a page from the results, open it to crawling and add noindex.

What is the maximum size of robots.txt?

Google reads the first 500 KiB of the file and ignores the rest. So long lists of rules are better shortened with wildcard patterns.

How do I check whether a specific URL is blocked?

Enter the site, and in the second field the page address, and choose a bot. We will show whether crawling is allowed or forbidden, which rule applied and on which line of the file.

How much does a robots.txt check cost?

Nothing: the check is free both on this page and in your StayIndexed account. In your account you can save a list of sites, refresh the check in one click and see the history of robots.txt changes.

Track the robots.txt of your sites

A list of sites, robots.txt errors and change history — in your account, free of charge. Sign-up without a card.

Start for free