Indexability Checker
Enter a URL and see whether Google can actually index it — with the four signals that decide it (meta robots, X-Robots-Tag, canonical, robots.txt) cross-checked for the contradictions that quietly keep pages out of search.
Crawlability and indexability are not the same thing
A page can be crawlable but not indexable, and the reverse causes the trickiest bugs. robots.txt controls crawling — whether Googlebot is allowed to fetch the URL at all. The meta robots tag and X-Robots-Tag control indexing — whether a page Google did fetch is allowed into the index. Because the two live in different places, a URL can pass one gate and fail the other. This checker reads both layers for a single URL so you can see, in one place, which gate is closing — instead of guessing from Search Console’s “Crawled — currently not indexed” label.
How to fix a page Google won’t index
Match the fix to the signal the tool flags:
- Verdict: noindex — remove the
noindexfrom the meta tag or the X-Robots-Tag header, then request re-crawl. Check both; a headernoindexoverrides even a clean<head>. - Verdict: canonicalised elsewhere — confirm the canonical target is the version you actually want ranked. If not, point the canonical back at the page itself.
- Verdict: blocked by robots.txt — if you want the URL indexed, unblock the path so Google can read the page and its on-page signals.
Frequently asked questions
What decides whether a page gets indexed?
Four signals, and they have to agree: the meta robots tag, the X-Robots-Tag HTTP header, the canonical tag, and robots.txt. This tool reads all four for a URL and tells you the net result — plus, crucially, any conflicts between them (the “crawled, currently not indexed” trap).
Why is “blocked by robots.txt AND noindex” a problem?
It’s the most common silent indexing bug. If robots.txt blocks the URL, Google can’t crawl the page — so it never sees the noindex tag. The page can then stay in the index (often as a bare URL with no description). The fix is to allow crawling so the noindex is actually read, or rely on robots.txt alone without the noindex.
What does “canonicalised elsewhere” mean?
Your page’s canonical tag points to a different URL, telling Google “index that one instead of this.” That’s correct for duplicate or parameter URLs, but if it’s pointing somewhere unexpected, this URL won’t get indexed and its ranking signals flow to the canonical target. The tool flags it so you can confirm it’s intended.
Does this check robots.txt the way Google does?
It applies the same core rules Google uses: it finds the most specific user-agent group, matches your URL’s path against the Allow/Disallow rules with wildcard (*) and end-anchor ($) support, and lets the longest matching rule win (with Allow breaking ties). It’s a faithful approximation for the common cases.
Embed this free tool on your site
Add this indexability checker to your own blog or website — it stays free for your visitors, and the short credit line links back here. Just copy the snippet and paste it into your page’s HTML.
<iframe src="https://seoaiaudits.com/embed/indexability-checker" width="100%" height="700" style="border:1px solid #e5e7eb;border-radius:12px;max-width:760px;width:100%" loading="lazy" title="Indexability Checker"></iframe> <p style="font-size:13px;font-family:system-ui,sans-serif;margin-top:8px;color:#6b7280"> Free <a href="https://seoaiaudits.com/tools/indexability-checker" target="_blank" rel="noopener">Indexability Checker</a> by <a href="https://seoaiaudits.com" target="_blank" rel="noopener">SEO AI Audits</a> </p>
Tip: adjust height if your layout needs more or less room. Please keep the credit line so others can find the tool.
Want a different one? See every tool you can embed.
Related free tools
Instant SEO Snapshot
One-click SEO & AI-readiness grade for any page — HTTPS, indexability, title, meta, headings, content & more. Embeddable on your own site.
AI Bot Access Checker
What does your robots.txt say to OAI-SearchBot, PerplexityBot, Claude-SearchBot, GPTBot and the rest? Read with the real rules.
AI Model Recall Checker
Ask free AI models the unbranded questions your buyers type, and count how many name your brand unprompted — with each model and its training cutoff shown. It measures what those models remember from training, not a live check of ChatGPT, Claude or Perplexity; a built-in panel lets you spot-check the real engines by hand.
NLP Content Analyzer
Topics, entities, search intent & content gaps an NLP model sees in your page.
AI Retrieval Checker
AI-retrieval readiness for one page — an AI model’s read of how findable and quotable it is, with fixes.
AI Crawler Log Analyzer
Upload your access log to see which AI bots (GPTBot, ClaudeBot, PerplexityBot) actually crawled you — with IP verification where the operator publishes ranges. Runs in your browser.
Want the full picture?
This tool checks one thing. Run a complete, free SEO audit across 28 modules.
Run a free SEO audit