What Is Index Coverage?
Index Coverage is a report within Google Search Console that provides a detailed breakdown of how Google has processed each URL on your website. It separates pages into four primary states: Valid (indexed), Valid with warnings (indexed but with minor issues), Excluded (not indexed), and Error (could not be crawled or processed). Understanding these states is fundamental to diagnosing why specific pages do not appear in Google Search.
The Excluded category is where most diagnostic work happens. Within Excluded, Google specifies the exact reason a page was not indexed. Common reasons include: "Excluded by 'noindex' tag" (the page contains a noindex directive either in the meta robots tag or the HTTP header); "Blocked by robots.txt" (Googlebot was prevented from crawling the page by a robots.txt disallow rule, so it cannot be indexed); "Crawled, currently not indexed" (Google visited the page but chose not to index it, often because it judged the content to be low quality or thin); "Duplicate, submitted URL not selected as canonical" (the page has a canonical tag pointing to a different URL and Google agrees with that designation); "Page with redirect" (the URL redirects to another page); and "Soft 404" (the page returns a 200 status code but its content resembles a page not found or error state).
The distinction between "Crawled, currently not indexed" and all other Excluded states is particularly important. Most other excluded states have a clear technical fix. "Crawled, currently not indexed" is a quality signal: Google visited the page but made a discretionary decision not to include it in the index. The fix is not technical but editorial, meaning the page needs substantively better, more useful content before Google will consider indexing it.
Index Coverage was renamed "Indexing" in the updated Search Console interface in 2023, but the underlying data and categories remain the same. Most SEO practitioners still refer to it by its original name.
Index Coverage In Practice
A Pretoria-based legal services firm had a 180-page website and was concerned that their blog content was not appearing in search results despite six months of regular publishing. The Index Coverage report showed 72 pages in the Excluded state. Investigating the breakdown revealed two separate issues.
First, 48 pages were excluded with the reason "Duplicate, Google chose different canonical than user." The firm's WordPress theme had been generating both www and non-www versions of every URL without a consistent canonical tag, causing Google to treat half the site as duplicates. Setting a consistent canonical domain in the site settings resolved this in the subsequent crawl cycle.
Second, 24 blog posts carried the status "Crawled, currently not indexed." These were posts that had been published at under 400 words with minimal original analysis, covering topics that dozens of other legal blogs had covered in greater depth. The team rewrote 12 of the highest-priority posts, expanding them to over 900 words with specific South African legal references, case examples, and structured FAQs. Within eight weeks, 10 of the 12 rewritten posts had been added to the index and were generating impressions in the Performance report.
What index coverage shows
Index coverage refers to which of a site's pages are included in Google's index and which are not, and to the report in Google Search Console that shows this, along with the reasons pages are or are not indexed. Because only indexed pages can appear in search results, understanding index coverage is fundamental: it tells you whether the pages you want found are actually in Google's index, and if not, why. The Search Console coverage (or page indexing) report categorises pages as indexed or not indexed, and for those not indexed, gives reasons, such as excluded by a noindex tag, blocked by robots.txt, marked as a duplicate with a canonical elsewhere, crawled but not indexed (often a quality judgement), discovered but not yet crawled, or various errors. This makes the report a primary diagnostic tool: it surfaces indexing problems, pages you wanted indexed that are not, and lets you address the specific cause, which is essential because a page that is not indexed cannot rank no matter how good it is.
Improving index coverage
Improving index coverage means getting the pages you want indexed into Google's index by addressing whatever is keeping them out, which the coverage report helps identify. The approach depends on the reason. Pages excluded by a noindex tag or blocked in robots.txt when you actually want them indexed need those directives corrected. Pages marked as duplicates need the canonical situation resolved so the right version is indexed. Pages crawled but not indexed are often a quality signal, Google chose not to include them, so improving thin or low-value content, or consolidating it, is the remedy. Pages discovered but not yet crawled may need better internal linking and inclusion in the sitemap so Google prioritises them, and on large sites, crawl efficiency matters. Ensuring important pages are well-linked internally, included in an accurate XML sitemap, free of accidental noindex or canonical errors, and genuinely worth indexing addresses most coverage problems. Because being indexed is the precondition for ranking, monitoring index coverage and resolving why valuable pages are excluded is a basic, high-priority part of technical SEO, since perfecting a page that Google will not index achieves nothing.
FAQ
Why are some of my pages excluded from Google's index?
The most common reasons include: a noindex tag on the page (intentional or accidental), the page being blocked by robots.txt (preventing crawling), a canonical tag pointing to a different URL, the page returning a redirect or error status code, Google detecting the page as a near-duplicate of another URL, or the page being identified as a soft 404 (returning a 200 status code but containing very little content). The Index Coverage report in Google Search Console lists the specific reason for each excluded URL.
How long does it take for a new page to appear in Google's index?
Indexing time varies widely. For sites with regular crawl activity, new pages on established domains often appear within a few days. For newer sites or pages with few internal links, it can take weeks or longer. Submitting the URL via the URL Inspection tool in Google Search Console and including the page in your XML sitemap can speed up discovery. Pages from established South African businesses on active domains typically index within 3 to 7 days.
Why are some pages excluded from Google's index?
For various reasons the coverage report specifies: a noindex tag or robots.txt block, being marked a duplicate with the canonical elsewhere, being crawled but judged not worth indexing (a quality signal), being discovered but not yet crawled, or errors. The remedy depends on the reason, correcting directives, resolving duplication, improving thin content, or strengthening internal links and the sitemap.