Page indexing report · Search Console

Blocked by robots.txt

A rule in your robots.txt file stops Google from crawling the URL. Google’s Help page calls this “URL blocked by robots.txt”.1 For carts, accounts and search pages, that’s the point. For pages you want in search, it’s a rule to find and change, and the tester on this page shows you which one.

Updated · Based on Google’s documentation, statements from Google staff and published studies · 12 sources

Check the URL first

The checker fetches your live robots.txt and tells you whether it blocks Googlebot from this URL. Google may still be using a copy from up to a day ago, and the report shows its last check.

Free, no sign-up. Checks status, noindex, robots.txt, canonical, content and sitemap.

In 30 seconds

  • robots.txt controls crawling, not indexing. It’s not a way to keep a page out of Google.2
  • A blocked URL can still be indexed from links, without its content.1
  • Google can’t see a noindex on a page it isn’t allowed to crawl.5
  • Google caches robots.txt for up to 24 hours, so fixes aren’t instant.3

What the status means

Before Googlebot fetches a URL, it checks the site’s robots.txt. If a rule disallows the URL, Google doesn’t request it at all. The report lists it here and notes that this doesn’t guarantee the page stays out of the index: if Google finds information about it elsewhere, there’s a small chance it gets indexed anyway.1

This is the one status in this group that’s set at the crawl stage, before Google has seen anything on the page. Select a stage to see what happens there.

Crawl

Googlebot checks robots.txt, then requests the URL. It spaces requests out so it doesn’t overload the server.

What goes wrong: The URL is blocked, returns an error, or the server is slow, so Google backs off and crawls less.

Report statuses at this stage

Google’s John Mueller has described robots.txt as close to absolute: if Google can parse the file, it follows the rules.9 So this status is never Google being unreliable. A rule matched.

Is it a problem?

Usually fine

  • Cart, checkout, account and order pages
  • Internal search results
  • Sorting, filter and session parameter URLs
  • Admin areas and API endpoints
  • Platform defaults, such as Shopify’s

Worth fixing

  • Pages you want in search: products, articles, landing pages
  • The whole site, after a launch or migration
  • CSS, JavaScript or API files that your pages need to render
  • Pages you blocked to get them out of Google
  • Blocked URLs listed in your sitemap

Blocked page resources matter more than they look. If Google can’t load scripts or styles a page needs, it can get a blank or broken page, which can lead to a soft 404.7

Find the cause

Answer a few questions. Each cause is explained below.

Question 1What do you want Google to do with this URL?

A deliberate block

robots.txt exists to manage which URLs crawlers request, mainly to avoid overloading the site.2 Blocking carts, accounts and infinite filter combinations is normal. Shopify’s default file blocks admin, cart, account, order and sorted collection pages for exactly that reason.11

A staging file or site-wide setting went live

A Disallow: / meant for a test site is the classic launch-day mistake. Older WordPress versions added it when “Discourage search engines” was ticked. Since version 5.3, that setting uses a noindex tag instead.12

A rule that matches more than you meant

Rules match from the start of the path, paths are case-sensitive, and * and $ are wildcards. When rules conflict, the longest matching path wins.3 So Disallow: /p blocks /products, and Disallow: /Blog/ doesn’t block /blog/.

Blocking pages to remove them

Google says robots.txt is not a mechanism for keeping a page out of Google. A blocked URL can still appear, without a description.2 And if the page carries a noindex, Google never sees it, because it can’t crawl the page.5 A Noindex: line in robots.txt doesn’t help: Mueller has called it an unsupported directive that does nothing.10

robots.txt is unreachable

If the file returns a server error, Google pauses crawling for 12 hours and then relies on its cached copy for up to 30 days.3 Firewalls and bot protection that block Googlebot from the file can do the same. Check the robots.txt report in Search Console for fetch errors.4

Test your robots.txt

Paste your robots.txt and a URL path to see which rule applies, using Google’s matching rules. Copy the file from https://yoursite.com/robots.txt.

Blocked

Matched Disallow: /*?sort= in the User-agent: * group. It’s the longest rule that matches.

Runs in your browser with Google’s matching rules: the most specific user-agent group applies, the longest matching path wins, and Allow wins a tie.

How Google reads robots.txt

The rules most often behind a surprise block. Pick one for details.

Caching

Google generally caches robots.txt for up to 24 hours, longer if it can’t fetch a fresh copy. A Cache-Control: max-age header can change that.3 To speed things up after a fix, use Request a recrawl in Search Console’s robots.txt report.4

How to fix it

  1. Decide whether each blocked URL should be crawled. Most sites have some that shouldn’t.
  2. For the ones that should, find the matching rule with the tester above. Make sure you’re reading the robots.txt for the right host and protocol.
  3. Remove or narrow the rule, or add a longer Allow rule for the path you want crawled.3
  4. In Search Console, open the robots.txt report and request a recrawl of the file.4
  5. Request indexing for your most important URLs, and validate the fix in the report.
  6. For pages you blocked to remove them: unblock them and add noindex instead, or return 404 or 410.6

What other guides get wrong

Several popular guides, including some from 2025, still give this advice:

Often saidWhat’s actually true
Check the rule with the robots.txt Tester in Search ConsoleThat tool has been replaced by the robots.txt report, which shows the files Google found, when it crawled them, any errors, and lets you request a recrawl.4 Google’s Help page for this status still mentions the tester.

How long it takes

Google gives no timeline for recrawling and indexing the unblocked pages themselves.

Blocked vs Indexed, though blocked by robots.txt

Blocked by robots.txtIndexed, though blocked by robots.txt
Crawled?NoNo
Indexed?NoYes, from links to it, without reading the page1
Report sectionNot indexedIndexed, with a warning
If you want it out of GoogleUnblock it and add noindexUnblock it and add noindex1

Questions

What does “Blocked by robots.txt” mean in Search Console?

A rule in your robots.txt file stops Googlebot from crawling the URL, so Google can’t read the page. Google’s Help page calls it “URL blocked by robots.txt”. It’s fine for pages you don’t want crawled.

What is the difference between Blocked by robots.txt and Indexed, though blocked by robots.txt?

Both mean robots.txt stops Google crawling the URL. In the first, the URL isn’t indexed. In the second, Google indexed the URL anyway, from links pointing to it, without reading the page.

How do I unblock a page in robots.txt?

Find the Disallow rule that matches the URL and remove or narrow it, or add a longer Allow rule. Then request a recrawl of robots.txt in Search Console’s robots.txt report and request indexing for the page.

Should I use robots.txt to remove pages from Google?

No. Google says robots.txt doesn’t keep pages out of Google, and a blocked page can still be indexed from links. Use a noindex tag, a login or a 404 or 410 instead, and leave the page crawlable so Google sees it.

Where is the robots.txt tester in Search Console?

The old robots.txt Tester was replaced by the robots.txt report. It shows the robots.txt files Google found, when it last crawled them, and any errors, and lets you request a recrawl. To test a URL against your rules, use the tester on this page.

Why is my Shopify store blocked by robots.txt?

Shopify’s default robots.txt blocks admin, cart, account, order and sorted collection pages, which shouldn’t be in search anyway. If a page you want indexed is blocked, you can edit the robots.txt.liquid template, though Shopify Support doesn’t help with those edits.

Sources

  1. Google, “Page indexing report”, Search Console Help
  2. Google, “Introduction to robots.txt”, Search Central, updated December 2025
  3. Google, “How Google interprets the robots.txt specification”, Search Central, updated August 2026
  4. Google, “robots.txt report”, Search Console Help
  5. Google, “Robots meta tag, data-nosnippet, and X-Robots-Tag specifications”, Search Central, updated March 2026
  6. Google, “Block Search indexing with noindex”, Search Central, updated December 2025
  7. Google, “Troubleshoot Google Search crawling errors”, Search Central, updated December 2025
  8. Google, “Removals and SafeSearch reports tool”, Search Console Help
  9. John Mueller (Google), LinkedIn, September 2024, reported by PPC Land
  10. John Mueller (Google), LinkedIn, March 2024, reported by Search Engine Journal
  11. Shopify, “Editing robots.txt.liquid”, Shopify Help Center
  12. Peter Wilson, “Changes to prevent search engines indexing sites”, Make WordPress Core, September 2019