Back to guides

Page Indexability: Why Search Engines Skip Pages You Want Ranked

How noindex tags, X-Robots-Tag headers and robots.txt rules stop pages being indexed, and how to confirm the pages you care about are open.

Indexability is the first gate. A page search engines decline to index cannot rank, cannot earn a snippet, and cannot appear in an AI answer, however good the writing is. The directives that control it are small and easy to move by accident: a noindex meta tag, an X-Robots-Tag header, a robots.txt rule, or a canonical pointing somewhere else. Most indexability damage traces back to a deploy. A staging build ships with its blanket noindex intact. A CDN rule adds X-Robots-Tag: noindex across a whole path. A CMS setting flips for one template and quietly removes 200 product pages. All of it returns a healthy 200, so uptime monitoring stays green while traffic drains over weeks. Read the directives the way a crawler does, in the two places that often disagree: the HTML you serve and the headers wrapped around it. A sitemap that lists blocked or noindexed URLs asks search engines to index pages you have simultaneously told them to skip; the XML sitemaps guide covers that conflict, and it is worth reading next. Work through these in order. Confirm the homepage is open, confirm the header layer agrees with the HTML, then confirm the pages you actually care about are reachable.

Uptime monitoring cannot see this kind of drift, but a weekly re-audit can. Paid plans include monitoring: register a domain once and it is re-audited every week, with an email when something important changes. Details: SEO monitoring API.

SEOReport's paid diagnosis reviews this across the pages of your own site, shows the evidence behind every finding, and ranks the fixes by priority. See plans and pricing.

Get the complete diagnosis of your site

An evidence-backed report and a prioritized action plan, on a plan with monthly credits.