Why Is Google Not Indexing All My Pages?
On this page
When a large site is mostly indexed but a slice stays out under two statuses, that slice is less a malfunction to force open than a mix of page-by-page choices and crawl scheduling. Google’s help for the Page indexing report says it plainly: “Not indexed” is not necessarily bad, and it’s fine for a URL not to be indexed for the right reasons. On a site with thousands of templated pages, a useful question is what separates the pages Google kept from the ones it left out.
First, split the unindexed pages by status
Two statuses in the report call for different work:
- Crawled – currently not indexed. Google fetched the page and didn’t index it. The report’s help says such a page may or may not be indexed in the future, and that there’s no need to resubmit it for crawling. Sending the same page back unchanged is not the lever; changing the page is the lever you control.
- Discovered – currently not indexed. Google found the URL but hasn’t crawled it yet. Google’s explanation is capacity: typically it wanted to crawl the URL, expected that to overload the site, and rescheduled.
The first is a question about the pages. The second is a question about crawling: how many wasted URLs compete for it and how much load your server can take. Start by cutting the waste that Google’s crawl budget guide names, such as duplicate URLs and soft 404s.
Let the indexed pages point to the threshold
A clear case is a directory or listing site, where every page is built from the same structured fields: name, address, category, hours. If a dozen larger sources publish those same fields for the same venue, your version gives Google little it doesn’t already have.
The pages Google did index are the evidence of what it is keeping. The indexed set is a guide to the threshold, and the difference between the indexed and unindexed pages of one template can point to the differentiator you are missing at scale. Read it directly:
- Pick one template. Comparing a product page with a blog post tells you little.
- Take a sample from each side. The report lists example URLs for indexed pages and for each not-indexed reason. Its help says those lists are capped at 1,000 rows and don’t necessarily show every URL, so treat them as samples. For example, take 20 indexed and 20 “Crawled – currently not indexed” URLs from the template.
- Record the same fields for all 40. Original description or not, reviews, photos, a unique data point, an editorial note, how many internal links point in.
- Find what the indexed side has and the other side lacks. That is your working threshold.
Treat the result as a hypothesis, not a rule Google stated. Test it: add the differentiator to one batch of unindexed pages, leave a similar batch unchanged, and compare after Google has recrawled both.
What doesn’t move the number
- Requesting indexing for unchanged pages. The URL Inspection tool documentation says submitting a request does not guarantee that a page will appear in the index. For many new or updated pages, the same documentation points to a sitemap with the updated pages marked by
lastmod. - Publishing more pages from the same fields. Each new page that repeats what the unindexed ones already repeat can add to the gap instead of closing it.
- Adding
noindexto free up crawling. It keeps a page out of results, but the crawl budget guide says Google still requests the page and then drops it, which spends crawling time rather than saving it. Use it for pages people need but searchers don’t, not as a crawl lever.
Decide which pages are meant to be indexed
Stop counting submitted pages and decide, per template, which pages are supposed to be in the index. For example, a site with a thousand pages that each carry something distinct, 950 of them indexed, is in a better position than one that submitted fifteen thousand and has four thousand indexed. The second site’s problem may be less indexing than a publishing decision to correct.
Then act on what the comparison showed. If only pages with reviews get indexed, either earn that differentiator wherever you can, or accept that the rest of the template doesn’t belong in the index, and stop generating pages that can’t clear it. Partial indexing on a large site can be Google telling you, page by page, where your value is. The work is to agree with that verdict where it’s right and change the pages where you can.
Frequently asked questions
Will requesting indexing fix “Crawled – currently not indexed”?
Not for an unchanged page. Google’s help says there’s no need to resubmit such a page, and that a request doesn’t guarantee indexing. Change the page first, then let Google recrawl it.
Should I noindex the pages Google won’t index?
Only if they still serve visitors and you want them out of search on purpose. noindex doesn’t save crawling, because Google still requests the page. Pages that serve neither visitors nor searchers are candidates for removal or consolidation instead.
How do I find out what my indexed pages have in common?
Compare samples from one template: indexed URLs against “Crawled – currently not indexed” URLs, field by field. The fields that appear on one side and not the other are your working threshold, to be tested on a batch before you roll it out.