When Good Pages Vanish From Google
A SaaS company we worked with published 40 new feature pages over a quarter, then watched their traffic stay flat. Every page loaded fine. None of them had a typo in the robots.txt. Yet when their team opened the Page Indexing report in Google Search Console, two thirds of those URLs sat under a status called “Crawled – currently not indexed.” Google had visited the pages, read them, and decided they were not worth adding to the index. No error message, no warning, just silence.
That scenario is more common than most marketing leaders realize, and it has gotten more common since 2025. If you want to fix Google indexing issues, the first thing to understand is that “not indexed” is rarely a bug. It is usually a judgment. This guide walks through the Page Indexing report state by state, explains what each status is telling you, and gives you the specific fixes that move pages from excluded back into search.
Start In The Right Report
Open Google Search Console, select your property, and go to Indexing, then Pages. This is the Page Indexing report, and it is the single source of truth for what Google has and has not added to its index.
At the top you see two buckets: pages that are indexed and pages that are not indexed. The number that matters is the gap between “discovered” pages and indexed pages, and the table beneath the chart that breaks down every reason a URL was left out. Click any row in that table and you get the exact list of affected URLs, which you can export.
Before you touch anything, confirm you are looking at the right property type. A domain property covers every subdomain and protocol, while a URL-prefix property only covers one exact prefix. Diagnosing indexing problems on the wrong property is the most common reason teams chase ghosts. A thorough SEO and GEO audit always begins by reconciling property coverage against the live sitemap so the numbers actually mean something.
Decode The Two States That Matter Most
Most indexing headaches come down to two statuses, and people constantly confuse them. The difference is not cosmetic. It tells you whether your problem is discovery and crawling, or quality and selection.
Discovered – Currently Not Indexed
Google’s own definition is precise: “The page was found by Google, but not crawled yet. Typically, Google wanted to crawl the URL but this was expected to overload the site; therefore Google rescheduled the crawl.” In plain terms, Google knows the URL exists, usually from your sitemap or an internal link, but it has not fetched the content. The page is in line, and the line is not moving.
This is fundamentally a crawl and priority problem. Google is rationing the time it spends on your site, a concept it calls crawl budget. According to Google Search Central, crawl budget is set by two forces: a crawl capacity limit, which is the maximum number of simultaneous connections Google will open without straining your server, and crawl demand, which reflects how much Google actually wants your content based on site size, freshness, and quality.
Crawled – Currently Not Indexed
Here Google states: “The page was crawled by Google but not indexed. It may or may not be indexed in the future; no need to resubmit this URL for crawling.” Google fetched the page, read it, and chose not to include it. This is not a crawl problem. It is an evaluation problem, and it is almost always about content quality, duplication, or thinness.
The distinction is the whole game. Discovered means “we have not looked yet.” Crawled means “we looked and we are not impressed.” You fix them in completely different ways, and treating a quality problem like a crawl problem (or the reverse) wastes weeks.
Why This Got Harder In 2025 And 2026
Google raised the bar for what earns a spot in the index, and the data shows it. One analysis of 1.7 million URLs found that roughly 88 percent of non-indexed pages were excluded for quality reasons rather than technical errors. After Google’s quality reviews through 2025, domains leaning on mass-produced, unedited AI content saw crawl priority fall sharply, and pages that did not demonstrate genuine experience and expertise were quietly left out.
There is a second reason indexing now carries higher stakes. As AI-driven answers and AI Overviews appear across a growing share of searches, a page that is not indexed cannot be cited, surfaced, or summarized in any of those experiences. Being absent from the index now means being invisible to both classic search and the AI answer layer, which is why getting indexing right is foundational to any serious AI search optimization (GEO) strategy.
The Step By Step Fix Process
Work through these steps in order. The sequence matters, because there is no point improving content on a page Google cannot reach, and no point requesting indexing on a page Google will reject.
Step 1: Confirm The Page Is Actually Indexable
Take a sample of affected URLs and run each through the URL Inspection tool at the top of Search Console. Use the live test, not just the indexed version, so you see how the page responds right now. Check three things: the page returns a 200 status code, it is not blocked in robots.txt, and it carries no noindex directive in either the meta tag or the X-Robots-Tag HTTP header.
A misconfigured noindex is the single most common “invisible” cause, and it appears in the report under its own status, “URL marked noindex,” defined by Google as “When Google tried to index the page it encountered a noindex directive and therefore did not index it.” If you find one on a page that should rank, remove it. This kind of header-level misconfiguration is exactly what technical SEO work catches and resolves at scale.
Step 2: Check For Canonical And Duplicate Conflicts
If your URLs land under “Duplicate without user-selected canonical” or “Duplicate, Google chose different canonical than user,” Google has decided your page is a copy of another and is serving the other version instead. Google describes this as a page that “is a duplicate of another page” where “Google has chosen the other page as the canonical.”
Audit the affected templates for near-identical content, parameter variations of the same page, and self-referencing canonical tags that point somewhere unexpected. Set a clear, self-referencing canonical on the version you want indexed, and make sure your internal links and sitemap point to that same canonical URL. Mixed signals here are a leading cause of pages being collapsed into a single indexed version.
Step 3: Fix Crawl Priority For Discovered Pages
If your problem is “Discovered – currently not indexed,” focus on making Google want to crawl, and able to crawl efficiently. The highest-leverage moves are:
- Add strong internal links from pages that already rank well to the orphaned or weakly linked URLs. A page buried five clicks from the homepage with no contextual inbound links signals low importance.
- Clean your XML sitemap so it lists only canonical URLs that return a 200 status. Remove redirects, error pages, and noindexed URLs, because a sitemap full of junk teaches Google to trust it less.
- Eliminate redirect chains and crawl traps. Google explicitly warns that long redirect chains “have a negative effect on crawling,” and faceted navigation or endless parameter URLs can drain crawl budget. Block genuine crawl traps in robots.txt rather than relying on noindex, which still forces a fetch.
- Improve server response time. The Crawl Stats report (Settings, then Crawl stats) shows whether slow responses or 5xx errors are throttling Googlebot. A faster, more reliable server lifts the crawl capacity limit.
Most of these levers only matter at scale. Google is clear that crawl budget is mainly a concern for sites with more than a million pages updated weekly, sites of 10,000-plus pages updated daily, or any site with a large share of URLs stuck in “Discovered – currently not indexed.” For a 200-page marketing site, the fix is almost never crawl budget and almost always internal linking and quality.
Step 4: Raise Quality For Crawled But Not Indexed Pages
When pages are “Crawled – currently not indexed,” Google has already read them and passed. As Yoast frames it, the question you have to answer is blunt: why should Google even index this page? Improving these pages means giving Google a reason to reverse its decision.
Look for thin content that restates what dozens of other pages already say, doorway-style pages that exist only to target a keyword, and templated pages where 90 percent of the copy is boilerplate. Consolidate weak pages into stronger, more comprehensive ones. Add original data, firsthand experience, expert commentary, and the kind of specificity a generic page cannot fake. This is where deliberate content marketing strategy and careful on-page optimization turn a skipped URL into an indexed, ranking asset.
Step 5: Handle Soft 404s And Server Errors
If you see “Soft 404,” Google is telling you a page “returns a user-friendly not found message but not a 404 HTTP response code.” Thin category pages, empty search results, and out-of-stock product pages frequently trigger this. Either build the page out into something substantive or return a proper 404 or 410 for genuinely dead URLs. For “Server error (5xx)” rows, work with your engineering team to resolve the underlying instability, because repeated 5xx responses make Google back off crawling entirely.
Step 6: Validate And Request Indexing
Only after the root cause is fixed should you ask Google to recheck. In the report, open the relevant status and click Validate fix. Google then recrawls the affected URLs and tracks progress, a process that typically takes up to two weeks and ends with an email notification. For individual high-value pages, use the URL Inspection tool and select Request indexing. Do not request indexing on pages you have not actually improved, because Google will simply crawl them again and reach the same conclusion.
How To Tell Whether It Worked
Indexing fixes are not instant. Plan for a two to four week window for changes to take effect, and longer for systemic, site-wide quality problems. Track the affected URL count in the Page Indexing report weekly, not daily, so you are reading a trend rather than noise. The right signal is a steady decline in the not-indexed bucket and a matching rise in indexed pages, ideally accompanied by impressions returning in the Performance report.
If you manage a large or fast-changing site, build this into a standing process rather than a one-time cleanup. Ongoing reporting and analytics that watch index coverage, crawl stats, and impressions together will catch indexing regressions weeks before they show up as a traffic drop, which is exactly when they are cheapest to fix.
The Takeaway For Decision Makers
Indexing is not a vanity metric. A page that is not in the index earns nothing in search and cannot appear in AI answers, no matter how much you spent producing it. The good news is that the Page Indexing report tells you precisely what is wrong if you read it correctly: discovered means crawl and priority, crawled means quality and selection, and each has a distinct fix. Diagnose the right state, address the root cause, validate, and the index follows. The teams that treat this as a recurring discipline, not a fire drill, are the ones whose new pages start ranking on schedule instead of disappearing into silence.
Sources
See where you are cited today
A free snapshot audit of your rankings and AI citations before we ever talk.