Indexability Checker

Enter a URL and see whether Google can actually index it — with the four signals that decide it (meta robots, X-Robots-Tag, canonical, robots.txt) cross-checked for the contradictions that quietly keep pages out of search.

Crawlability and indexability are not the same thing

A page can be crawlable but not indexable, and the reverse causes the trickiest bugs. robots.txt controls crawling — whether Googlebot is allowed to fetch the URL at all. The meta robots tag and X-Robots-Tag control indexing — whether a page Google did fetch is allowed into the index. Because the two live in different places, a URL can pass one gate and fail the other. This checker reads both layers for a single URL so you can see, in one place, which gate is closing — instead of guessing from Search Console’s “Crawled — currently not indexed” label.

How to fix a page Google won’t index

Match the fix to the signal the tool flags:

  • Verdict: noindex — remove the noindex from the meta tag or the X-Robots-Tag header, then request re-crawl. Check both; a header noindex overrides even a clean <head>.
  • Verdict: canonicalised elsewhere — confirm the canonical target is the version you actually want ranked. If not, point the canonical back at the page itself.
  • Verdict: blocked by robots.txt — if you want the URL indexed, unblock the path so Google can read the page and its on-page signals.

Frequently asked questions

What decides whether a page gets indexed?

Four signals, and they have to agree: the meta robots tag, the X-Robots-Tag HTTP header, the canonical tag, and robots.txt. This tool reads all four for a URL and tells you the net result — plus, crucially, any conflicts between them (the “crawled, currently not indexed” trap).

Why is “blocked by robots.txt AND noindex” a problem?

It’s the most common silent indexing bug. If robots.txt blocks the URL, Google can’t crawl the page — so it never sees the noindex tag. The page can then stay in the index (often as a bare URL with no description). The fix is to allow crawling so the noindex is actually read, or rely on robots.txt alone without the noindex.

What does “canonicalised elsewhere” mean?

Your page’s canonical tag points to a different URL, telling Google “index that one instead of this.” That’s correct for duplicate or parameter URLs, but if it’s pointing somewhere unexpected, this URL won’t get indexed and its ranking signals flow to the canonical target. The tool flags it so you can confirm it’s intended.

Does this check robots.txt the way Google does?

It applies the same core rules Google uses: it finds the most specific user-agent group, matches your URL’s path against the Allow/Disallow rules with wildcard (*) and end-anchor ($) support, and lets the longest matching rule win (with Allow breaking ties). It’s a faithful approximation for the common cases.

Embed this free tool on your site

Add this indexability checker to your own blog or website — it stays free for your visitors, and the short credit line links back here. Just copy the snippet and paste it into your page’s HTML.

Embed code (HTML)
<iframe src="https://seoaiaudits.com/embed/indexability-checker" width="100%" height="700" style="border:1px solid #e5e7eb;border-radius:12px;max-width:760px;width:100%" loading="lazy" title="Indexability Checker"></iframe>
<p style="font-size:13px;font-family:system-ui,sans-serif;margin-top:8px;color:#6b7280">
  Free <a href="https://seoaiaudits.com/tools/indexability-checker" target="_blank" rel="noopener">Indexability Checker</a> by <a href="https://seoaiaudits.com" target="_blank" rel="noopener">SEO AI Audits</a>
</p>

Tip: adjust height if your layout needs more or less room. Please keep the credit line so others can find the tool.

Want a different one? See every tool you can embed.

Related free tools

Want the full picture?

This tool checks one thing. Run a complete, free SEO audit across 28 modules.

Run a free SEO audit