The Page indexing report in Google Search Console (found under Indexing › Pages) is where most website owners first discover that something is wrong. It shows how many of your URLs Google has indexed, and lists the reasons the rest are not indexed. The problem is that the report mixes harmless housekeeping with serious issues, all under the same grey heading. This guide translates every status into plain English and tells you which ones matter.
This article is part of our complete guide: Why Is My Website Not on Google? The Complete Indexing Guide.
How the report is structured
At the top you will see a chart of indexed and not indexed pages over time. Below it, “Why pages aren’t indexed” lists each reason, the source (Website or Google systems) and the number of affected pages. Click any row to see example URLs and a chart for that reason.
Two tips before you start. First, filter the report by sitemap (“All submitted pages” in the drop-down at the top) to focus on URLs you care about. Second, remember the data is not real-time; it reflects Google’s last crawl of each URL.
Statuses that are usually normal
Alternate page with proper canonical tag
A duplicate version that correctly points to the main version. Usually healthy. Full guide.
Page with redirect
The URL redirects elsewhere, so Google indexes the destination. Normal for http, www and old URLs, but your sitemap and internal links should not point to redirects. Full guide.
Not found (404)
The page does not exist. Normal for deleted content and mistyped links. Only act if important pages or pages with backlinks return 404, in which case redirect them to a relevant page.
Excluded by ‘noindex’ tag
Normal for thank-you pages, carts, account pages and thin archives. Serious if your key pages are in the list. Full guide.
Blocked by robots.txt
Normal for admin and checkout areas. Serious if important content or resources are blocked. Full guide.
Statuses that usually need attention
Crawled – currently not indexed
Google read the page and decided not to index it, usually because of quality, duplication or low importance. This is one of the most important statuses for content-heavy sites. Full guide.
Discovered – currently not indexed
Google knows the URL but has not crawled it, usually because of crawl budget, slow servers or weak internal links. Full guide.
Duplicate without user-selected canonical
Google found duplicates and chose a canonical itself because you did not specify one. Check that Google chose the right version. Full guide.
Duplicate, Google chose different canonical than user
You declared a canonical, but Google disagreed. This often means your signals conflict: internal links, sitemaps or redirects point elsewhere, or the pages are not really duplicates. Align your signals and make the preferred page clearly the strongest.
Soft 404
The page returns 200 but looks empty or like an error. Add content, return a proper 404, or redirect to a relevant page. Full guide.
Server error (5xx)
Google could not load the page because the server failed. Occasional errors happen, but a growing number points to hosting problems, plugin conflicts or a firewall blocking Googlebot. Check with your host and review Crawl stats.
Redirect error
The redirect is broken: a loop, a chain that is too long, a redirect to an invalid URL or an extremely long URL. Fix the redirect so it goes directly to a working page.
Blocked due to unauthorized request (401) and access forbidden (403)
Google was asked for a login or refused access. Normal for private areas. If public pages are affected, a security plugin, firewall or CDN bot protection may be blocking Googlebot. Whitelist verified Googlebot rather than turning protection off.
Blocked due to other 4xx issue
A less common client error such as 410 or 429. A 429 (too many requests) means your server or firewall is rate-limiting Google.
Indexed, though blocked by robots.txt
Google indexed a URL it could not crawl, from links alone. Decide whether the page should be indexed (remove the block) or not (allow crawling and add noindex).
The order we fix things in
- Anything affecting key pages: homepage, services, products, top articles. Check these first regardless of status.
- Server errors and redirect errors, because they affect crawling site-wide.
- Accidental noindex and robots.txt blocks, because they are quick, binary fixes.
- Canonical issues, especially where Google chose a different page.
- Soft 404s, by adding content, redirecting or returning a real 404.
- Discovered and crawled – not indexed, which take longer and involve content, internal links and site quality.
How to validate fixes
After fixing an issue, open its row and click Validate fix. Google will recrawl the affected URLs over the next few weeks and report “Passed” or “Failed”. Validation is not required for Google to recognise fixes, but it speeds up rechecking and gives you a clear record.
Do not chase zero
A healthy website always has some not-indexed URLs: redirects, alternates, 404s and noindexed utility pages. Trying to reach zero wastes time. Instead, focus on one question: are all the pages that bring customers indexed? If yes, the rest is housekeeping.
Checking the report regularly
We recommend a quick look every month, and a deeper review after any redesign, migration, plugin change or theme update. Sudden jumps in “not indexed” are often the first sign of a technical problem, such as a plugin adding noindex or a hack creating spam pages. For WordPress sites, that last case is covered in our malware removal service.
Real examples
Across the sites in our case studies, from local service businesses like OneFlytt to content sites like M10 News, this report has been the starting point for every indexing project. Reading it correctly is what separates a two-hour fix from weeks of guesswork.
Related guides and services
- Google not indexing all pages: why only part of your site is indexed.
- Request indexing with URL Inspection: how to use the URL Inspection tool properly.
- How long does Google take to index a page: realistic timelines and how to speed them up.
- Website not showing on Google: the full diagnosis for sites that are missing from search.
- Google indexing fix service: we diagnose and fix pages Google will not index.
Get expert help
If you would rather spend your time running your business, MIVAQ can handle this for you. We audit the site, fix the root causes, document every change and show you the before-and-after in Search Console, so you know exactly what improved and why. Start with our Google indexing fix service, browse real client results in our case studies, or contact us for a free, no-pressure review of your website.
Frequently asked questions
Why does the report show different numbers from a site: search?
The site: operator is an estimate and often inaccurate. Search Console data is more reliable.
How often is the report updated?
Usually every few days, but each URL reflects Google’s last crawl of it, which may be older.
What does “Source: Google systems” mean?
Google considers the reason something it decided, rather than something your site requested, and you usually cannot validate it directly.
Can I see the report for a single page?
Yes. Use the URL Inspection tool for any individual URL.