Bulk Noindex & Indexability Checker

Check up to 20 URLs for noindex, X-Robots-Tag, robots.txt blocks, canonicals and status codes.

What makes a page indexable?

A page is indexable when Google is allowed to crawl it and add it to search results: it returns HTTP 200, has no noindex directive in a meta tag or X-Robots-Tag header, is not blocked by robots.txt and its canonical tag points to itself.

The five checks

  1. Status code — only 200 pages can be indexed.
  2. Meta robots / googlebot — a noindex value removes the page from Google.
  3. X-Robots-Tag header — the same directive sent as an HTTP header (common for PDFs).
  4. robots.txt — checked for Googlebot; blocked pages cannot be crawled.
  5. Canonical — a canonical pointing to another URL asks Google to index that URL instead.

Common reasons pages disappear from Google

  • A staging "noindex" left on after launch.
  • A robots.txt rule that blocks a whole folder by mistake.
  • Canonical tags that all point to the home page.

Note: robots.txt blocks crawling, not indexing. A blocked page can still appear without a description if other sites link to it — use noindex (and allow crawling) to remove a page reliably.

Frequently asked questions

How do I remove a page from Google?

Add a noindex meta tag, make sure robots.txt does not block the page so Google can see the tag, and optionally use the removal tool in Search Console.

Does robots.txt Disallow stop indexing?

Not reliably. It stops crawling, but a URL can still be indexed from links. Use noindex instead.

Why does a page say "canonical points elsewhere"?

Its canonical tag names a different URL, which tells Google to index that URL instead of this one.