Technical SEO
Pages Crawled but Not Indexed: the SEO Problem Almost No One Checks
If you go into Google Search Console and check "Pages" (formerly "Coverage"), you'll very likely see a category called "Crawled - currently not indexed." Most site owners never check it because it's technically not an "error" — it doesn't break anything visually. But on large sites, this category can pile up thousands of pages, and that does carry a real cost.
What does it actually mean?
It means Google visited that page (crawled it), but decided, at its own discretion, not to include
it in its search index. It's different from a 404 error or a robots.txt block — the
page exists, it works, and Google saw it. It simply didn't consider it worth indexing, almost
always for one of these reasons:
- Duplicate or very similar content to another page already indexed on the same site.
- Low perceived value content — very thin pages, auto-generated ones, or pages with little unique text.
- Insufficient quality signals — no internal links pointing to it, no traffic, no accumulated authority.
Why it matters to your crawl budget
Google doesn't have infinite resources to crawl your site every day — it assigns a crawl budget proportional to your domain's size and authority. Every time that budget gets spent visiting pages that will never be indexed, it's budget that isn't being used to discover and re-index your new or updated content faster. On sites with thousands of pages — typically news outlets, large e-commerce stores, or very old blogs — this effect can be significant.
How to spot it (without paid tools)
- Go to Search Console → Pages → the "Why pages aren't indexed" section.
- Look specifically for "Crawled - currently not indexed" and "Discovered - currently not indexed."
- Review a sample of those URLs: are they filter pages, infinite pagination, duplicate content, or just old, low-value pages?
- Cross-reference that list against your Analytics: if, on top of not being indexed, they also get no meaningful direct or internal traffic, they're strong candidates to improve, consolidate, or remove.
What to do once you find it
There's no one-size-fits-all fix — it depends on why Google decided not to index them. The typical
options are: improve and expand the content if it has real potential, consolidate several similar
pages into one more complete page (avoiding keyword cannibalization), add internal links from
higher-authority pages, or explicitly mark with noindex the pages that should never
compete for rankings (like filter results or admin pages). This is exactly the kind of decision I
review as part of the 4F Method in every audit: cross-referencing the indexing report against real
Analytics performance before deciding what to do with each group of pages.
Not sure how many of your site's pages are in this situation?
I review your complete indexing status as part of every technical SEO audit.