The pages Google has never crawled, and why they can't be indexed
There are three reasons a page is not in Google. It is blocked. Its content is thin or duplicated. Or Google has never fetched it at all. The first two get most of the attention. The third is quieter, and on some sites it is the largest group.
Never crawled is not the same as not indexed
In Search Console, “Discovered – currently not indexed” means Google knows the URL exists but has not crawled it yet. Sitting behind that state is a set of URLs Google has no record of, and another set it crawled once a long time ago and has not returned to. We wrote about the crawled-not-indexed states separately. This post is about the pages before that point.
A page Google has never fetched has no content on file, no signals, and nothing to rank. It is not a bad page. It is an absent one. No amount of editing the page changes anything until Google actually requests it.
How much of your site is it
Every URL Inspection call returns the last time Google crawled that specific URL. Collect that across the whole site and you get a distribution: crawled this week, crawled this month, crawled months ago, never. Our Trends view shows this as a crawl-coverage bar so you can see the shape at a glance.
On a small, well-linked site the never-crawled slice is usually close to zero. It grows on large sites, on new sites, and on sites that recently changed their URL structure or navigation. If half your pages have no crawl on record, that is the number to work on first, ahead of any on-page tuning.
What puts a page in that bucket
A few things, often together:
- It is in no sitemap, or in a sitemap Google cannot fetch, or in one padded with redirects and
noindexURLs so Google trusts it less. - Nothing on the site links to it. Links are how Google discovers most pages, and a URL with no internal links is easy to miss. That case has its own post.
- Crawl budget. On a large site Google spends its crawling on the URLs it already believes are worth it. Thin, duplicated, and parameter-generated URLs eat the rest.
- It is buried. A page seven clicks from the homepage gets crawled rarely, if at all.
What actually moves it
Submit a clean sitemap in Search Console and confirm it parses with no errors. Keep it honest: only canonical, indexable URLs belong in it.
Add internal links from pages Google crawls often to the pages it does not. For most sites this moves the number more than anything else, and it is the one thing a sitemap alone cannot do.
For a small number of important URLs, use “Request indexing” in Search Console. It does not scale past a handful, but it works for those.
Cut the noise. Fewer thin and parameter URLs means more of Google’s crawling goes to the pages you care about. Then wait, and watch the last-crawl distribution shift over the following weeks.
Before any of that
Rule out a per-page blocker on the URLs you expect to be crawled. Our free checker tells you in a few seconds whether a specific URL is technically clear. If it is clear and Google still has no crawl on record, the problem is discovery, not the page. The product tracks the crawl-coverage distribution over time so you can tell whether your changes are working.