Discovered – Currently Not Indexed
What "Discovered – currently not indexed" means in Google Search Console, how it differs from "Crawled – currently not indexed," why it happens, and how to fix it.
1 evidence signal on this page
- Related live toolLog File Analyzer
"Discovered – currently not indexed" is a Google Search Console Page Indexing status: Google knows the URL exists (sitemap or link) but hasn't crawled it yet — the Last Crawl date is empty. That one fact separates it from "Crawled – currently not indexed," where the page was fetched and Google is still evaluating it for indexing. Google's two drivers are crawl capacity (crawling now would overload your server) and crawl demand (your site/pages aren't worth the crawl effort — a quality and internal-linking signal). It's often a sitewide signal, not a per-page bug, though that's a practitioner inference rather than a fact the status proves. The fixes are crawl-demand levers (internal links, content quality, cutting crawl waste, links to priority pages) and crawl-capacity levers (server speed/stability). "Request indexing" can nudge a few priority URLs but doesn't scale and doesn't fix the root cause — and getting a page crawled still doesn't guarantee indexing.
TL;DR — “Discovered – currently not indexedStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed.” in Google Search ConsoleA free Google service that reports how a site performs in Google Search and surfaces problems with how Google crawls, indexes, and serves it. It's first-party data straight from Google — but you don't need it to appear in results. means Google has foundA 302 (\"Found\") is a temporary redirect: it forwards users to a new URL while telling search engines the original URL should stay in the index. It's a weak canonicalization signal, not the zero-equity dead end of SEO folklore. your page but hasn’t downloaded (crawled) it yet — so it can’t be in search. The Last Crawl date is empty. It usually means Google didn’t think the page was worth crawlingCrawling is how search engines use automated bots (like Googlebot and Bingbot) to discover URLs and download pages. A page has to be crawlable to be indexed, but crawling on its own isn't a ranking factor. right now, or crawlingCrawling is how search engines use automated bots (like Googlebot and Bingbot) to discover URLs and download pages. A page has to be crawlable to be indexed, but crawling on its own isn't a ranking factor. it would have strained your server. The fix is to make your important pages easier to reach and clearly worth crawling — not to keep clicking “Request indexing.”
What this status means
Open Google Search ConsoleA free Google service that reports how a site performs in Google Search and surfaces problems with how Google crawls, indexes, and serves it. It's first-party data straight from Google — but you don't need it to appear in results., go to the Page IndexingThe Google Search Console report (formerly Index Coverage) showing how many of your URLs are indexed vs. not indexed, and grouping the not-indexed ones by reason. report, and you’ll see your pages grouped by status. “Discovered – currently not indexedA Google Search Console Page Indexing status meaning Google knows the URL exists but hasn't crawled it yet — the Last Crawl date is empty. Often a crawl-capacity or crawl-demand (site-quality) signal.” is the bucket for URLs that Google knows about but hasn’t actually crawled yet. Evidence for this claim Google defines Discovered - currently not indexed as a page it found but has not yet crawled, with an empty last crawl date. Scope: Google Search Console Page Indexing status. Confidence: high · Verified: Google: Page indexing report
Remember the three steps every page goes through to show up in search:
- Crawl — Google downloads the page.
- Index — Google files it away in its database.
- Serve (rank) — Google shows it when someone searches.
A “Discovered” page is stuck before step one. Google found the URL — usually from your sitemapA sitemap is a file that lists the pages, images, videos, and other files on your site so search engines can discover them. It helps discovery, but submitting a sitemap doesn't guarantee crawling or indexing. or a link — added it to its to-do list, and then didn’t get around to fetching it. The clearest sign is in the URL Inspection toolA Google Search Console feature that reports how Google sees one specific URL on a property you own. By default it shows the last-indexed snapshot; a separate \"Test live URL\" mode fetches the current version.: the Last Crawl date is empty, because nothing was ever crawled.
Why Google leaves a page “Discovered”
Two reasons, in plain terms:
- It didn’t want to overload your site. Google holds back if crawling more pages right now might slow your server down. It reschedules the crawl for later.
- It didn’t think the page was worth it. If your site (or that section of it) looks thin, duplicated, or hard to reach, Google deprioritizes crawling those URLs. This is a plausible diagnosis, not something the report status proves by itself.
How it’s different from “Crawled – currently not indexed”
These two statuses look almost identical and people mix them up constantly. The difference is one word — crawled:
- Discovered – Google hasn’t fetched the page yet. Empty Last Crawl date.
- Crawled – currently not indexedA Google Search Console Page Indexing status meaning Googlebot fetched the page but Google decided not to index it — usually a content- or site-quality signal, not a technical error. – Google did fetch it but decided not to keep it. There’s a Last Crawl date.
So Discovered is a “we haven’t gotten to it” problem; Crawled is a “we looked and passed” problem. Different stages, different fixes.
What actually helps
- Link to the page from pages that already get crawled — your homepage, main navigation, or popular articles. Orphan pagesAn orphan page is a page on your site that no other page links to internally. Because crawlers discover pages by following links, an orphan page is effectively invisible to search engines unless it's in an XML sitemap or linked from an external site. (nothing links to them) are one possible cause.
- Make the page genuinely useful and not a near-duplicate of others.
- Submit it in your XML sitemapAn XML sitemap is a UTF-8 file listing the canonical URLs on your site (with optional lastmod) so search engines can discover and prioritize them. It's a discovery and diagnostic aid, not a guarantee of indexing — and Google ignores its priority and changefreq tags. (this helps Google find it, though not necessarily prioritize it).
- Keep your server fast and stable.
The thing most people get wrong
Clicking “Request indexingStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed.” over and over is not the fix. It can nudge a few important URLs, but it doesn’t scale to hundreds or thousands of pages, and it does nothing about why Google deprioritized them. If a whole section is sitting in “Discovered,” that’s a signal about your site’s quality, structure, or server — not something a button solves. Evidence for this claim Google says repeated recrawl requests for the same URL do not make crawling faster and recommends sitemaps for many URLs. Scope: URL Inspection request indexing; crawling still does not guarantee indexing. Confidence: high · Verified: Google: Ask Google to recrawl URLs
Want the full diagnosis — how to tell whether it’s a server problem or a quality problem, and the fixes that actually scale? Switch to the Advanced tab.
TL;DR — “Discovered – currently not indexedStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed.” means Google foundA 302 (\"Found\") is a temporary redirect: it forwards users to a new URL while telling search engines the original URL should stay in the index. It's a weak canonicalization signal, not the zero-equity dead end of SEO folklore. the URL but hasn’t crawled it — the Last Crawl date is empty, which is the single fact that separates it from “Crawled – currently not indexedA Google Search Console Page Indexing status meaning Googlebot fetched the page but Google decided not to index it — usually a content- or site-quality signal, not a technical error.” (fetched, then not kept). Google’s two drivers are crawl capacityThe number of URLs an engine will crawl in a timeframe. (crawlingCrawling is how search engines use automated bots (like Googlebot and Bingbot) to discover URLs and download pages. A page has to be crawlable to be indexed, but crawling on its own isn't a ranking factor. now would overload the server, so it rescheduled) and crawl demandCrawl demand is the 'want' side of crawl budget — how much a search engine wants to crawl a site or URL, driven by popularity, staleness, and perceived inventory (plus temporary spikes from site moves). It's distinct from crawl rate/capacity, the 'can' side. (your site/pages aren’t worth the crawl effort — a quality and internal-linking signal). It’s often a sitewide pattern, not a per-page bug. Fix the demand side first — internal linking, content quality, cutting crawl waste, links to priority pages — and the capacity side (server speed/stability) for large sites. Request indexing nudges a handful of URLs, doesn’t scale, and doesn’t fix the cause. And crawlingCrawling is how search engines use automated bots (like Googlebot and Bingbot) to discover URLs and download pages. A page has to be crawlable to be indexed, but crawling on its own isn't a ranking factor. a page still doesn’t guarantee it gets indexed.
What Google actually says it means
Straight from the Page Indexing reportThe Google Search Console report (formerly Index Coverage) showing how many of your URLs are indexed vs. not indexed, and grouping the not-indexed ones by reason. definition: “The page was found by Google, but not crawled yet. Typically, Google wanted to crawl the URL but this was expected to overload the site; therefore Google rescheduled the crawl. This is why the last crawl date is empty on the report.” Evidence for this claim Google defines Discovered - currently not indexed as a page it found but has not yet crawled, with an empty last crawl date. Scope: Google Search Console Page Indexing status. Confidence: high · Verified: Google: Page indexing report That last sentence is the whole tell — empty Last Crawl date = never fetched. If you inspect one of these URLs in GSCA free Google service that reports how a site performs in Google Search and surfaces problems with how Google crawls, indexes, and serves it. It's first-party data straight from Google — but you don't need it to appear in results., you’ll see no crawl recorded.
So this is a pre-crawl queue state. Nothing was indexed and removed; nothing was penalized. Google knows the URL exists — it came in through a sitemapA sitemap is a file that lists the pages, images, videos, and other files on your site so search engines can discover them. It helps discovery, but submitting a sitemap doesn't guarantee crawling or indexing., an internal link, or an external link — and it simply hasn’t fetched it.
Discovered vs Crawled – currently not indexed
This is the distinction worth getting exactly right, because the two statuses have opposite root causes and opposite fixes. The contrast table lives in the Cheat Sheets tab; the short version:
- Discovered – currently not indexedA Google Search Console Page Indexing status meaning Google knows the URL exists but hasn't crawled it yet — the Last Crawl date is empty. Often a crawl-capacity or crawl-demand (site-quality) signal. = not yet fetched. Empty Last Crawl date. It’s a crawl-priority / capacity signal — Google decided not to spend a crawl on it (yet).
- Crawled – currently not indexed = fetched and not kept. There’s a Last Crawl date. Google looked and, for now, chose not to index it — an indexingStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed. evaluation that can have several causes (duplication, thin contentThin content is web content that provides little or no value to users. Google's spam policies name it 'thin content with little or no added value' — and it's about value per page, not word count., canonicalizationHow search engines pick one canonical URL among duplicates and consolidate signals onto it. to another URL, and more), not a single “quality verdict.”
Here’s the part people miss: getting a page out of Discovered doesn’t mean it gets indexed. It can move into Crawled – currently not indexed and still sit there. Crawling is a gate, not a guarantee — the same way it works everywhere else in search. Evidence for this claim Google says repeated recrawl requests for the same URL do not make crawling faster and recommends sitemaps for many URLs. Scope: URL Inspection request indexing; crawling still does not guarantee indexing. Confidence: high · Verified: Google: Ask Google to recrawl URLs
Google knows the URL. On the highlighted Discovered currently not indexed branch, Google has not fetched it, the Last Crawl field is empty, and diagnosis focuses on crawl priority or capacity. On the Crawled currently not indexed branch, Google fetched the page but did not index it, the Last Crawl field has a date, and diagnosis focuses on index selection, page value, duplication, rendering, and conflicting signals.
© Patrick Stox LLC · CC BY 4.0 ·
Why Google leaves pages in “Discovered”
Google frames crawling as a budget made of two halves, and Discovered is the canonical symptom of a problem on one side or the other.
Crawl capacity — your server
Google calculates a crawl capacity limit: the maximum number of simultaneous
connections it’ll use on your site, tuned to how your server responds. From the
crawl-budget guide: “Google’s crawlersA crawler — also called a spider or bot — is an automated program that fetches web pages, extracts their links, and queues new URLs to visit. Search engines use crawlers to discover and download content for their index. calculate a crawl capacity limit, which is
the maximum number of simultaneous parallel connections that Google can use to crawl
a site,” and “if the site slows down or responds with server errors, the limit
goes down and Google crawls less.” Slow responses, timeouts, and 5xx errors all
throttle the crawl — and when capacity is the bottleneck, URLs pile up in
Discovered because there literally wasn’t room to fetch them.
Crawl demand — your site’s quality and structure
The other half is whether Google wants to crawl the URL. This is where most “Discovered” problems actually live. Google’s systems extrapolate crawl priority from URL patterns, internal linkingLinks between pages on the same site., and overall site quality. If a page is buried deep, orphaned, or looks like one more near-duplicate in a large low-value set, demand for it is weak and it stays in the queue.
Notably, Google’s large-site crawl-budget guide explicitly calls out this status: it says the guide applies to “Sites with a large portion of their total URLs classified by Search ConsoleGoogle's free tool for monitoring crawling, indexing, and search performance. as Discovered - currently not indexed.” That’s Google itself tying Discovered to a crawl-budget (capacity + demand) constraint. Google also names the guide’s audience: sites with millions of URLs, sites that change rapidly and have roughly 10,000-plus pages, and sites with a large share of Discovered URLs — but it explicitly calls those numbers rough classification estimates, not exact thresholds. If your site is well under that scale, treat the guide as background, not a sign that a hard capacity ceiling applies to you.
It’s often a sitewide signal, not a per-page bug
This is a useful mental shift for most cases, though it’s a pattern I’ve observed rather than a frequency Google publishes. Discovered rarely means “page X has a flaw” in isolation. More often Google has extrapolated, from your URL patterns and sitewide quality, that a whole category of your pages isn’t worth crawling aggressively — but that’s a practitioner inference from URL-pattern and template behavior, not a fact the status itself proves for any single page. John Mueller has made the point repeatedly that there are two main drivers behind this status: server capacity (Google held back to avoid overloading the site) and overall website quality (the systems don’t think the pages are worth the crawl effort). He’s also noted the real-world causes are broader than the help doc’s “overload” line — accidentally auto-generating too many URLs, poor internal linking, and the need to strengthen the site overall so important pages get prioritized. (These are paraphrased from his office-hours commentary, relayed through industry coverage — I haven’t pinned them to a verbatim transcript.)
Scale plays a role too. Gary Illyes has been widely quoted, via industry coverage of his podcast remarks, as saying somewhere around 90% of sites don’t need to think about crawl budgetThe number of URLs an engine will crawl in a timeframe. at all — but I haven’t independently verified that figure against the original recording, so treat it as a widely relayed approximation, not a confirmed stat. The directional implication still holds: on a small or mid-size site, a true crawl-capacity ceiling is unlikely, and a persistent Discovered backlog is more often a demand problem — quality, internal linking, or crawl waste — than a server wall. Confirm that with your own Crawl Stats and logs rather than assuming it from site size alone.
How to diagnose which cause you have
Before you fix anything, work out whether you’re capacity-bound or demand-bound:
- Capacity check. Look at GSC Crawl Stats (average response time, host
status, response-code breakdown) and your server logs for slow responses and
5xx/timeout spikes. If Google is clearly being throttled by your server, that’s a capacity problem. - Demand check. Look at internal-link depth (how many clicks from the homepage), orphan pagesAn orphan page is a page on your site that no other page links to internally. Because crawlers discover pages by following links, an orphan page is effectively invisible to search engines unless it's in an XML sitemap or linked from an external site. (nothing links to them), and sitewide quality (thin, duplicated, or auto-generated URL patterns). If your Discovered URLs are deep, orphaned, or part of a near-duplicate set, that’s a demand problem.
Don’t pick a fix off a general frequency (“most sites are X”) — decide from the evidence in front of you: your own URL-pattern grouping, server logs, Crawl Stats, internal-link counts, sitemapA sitemap is a file that lists the pages, images, videos, and other files on your site so search engines can discover them. It helps discovery, but submitting a sitemap doesn't guarantee crawling or indexing./inventory coverage, and how much each affected group actually matters to the business. For most small and mid-size sites the evidence tends to point to demand; for very large, ecommerce, or programmatic sites, it’s often both — but confirm it on your own data before committing to a fix.
Join sitemap, fresh crawl, and Search Console exports to surface mismatches for review with my free Indexation Reconciler Free
- Load the sitemap inventory, a current crawl export, and the Search Console URL export for the same scope.
- Use mismatches to find stale exports, newly blocked URLs, and URLs absent from the intended inventory.
- Use Crawl Stats, server logs, internal-link evidence, and URL Inspection to decide whether the discovered backlog is capacity- or demand-driven.
How to fix it
The levers, roughly in order of impact for most sites:
Strengthen internal linking and fix orphan pages
Internal linking is the most controllable demand lever you have. Pages that nothing links to, or that sit many clicks deep, dominate the Discovered bucket. Link your important URLs from pages Google already crawls often — the homepage, hub pages, main navigation — and pull them shallower in the architecture.
Improve content quality; consolidate thin and duplicate pages
If Google is reading “low value” off your URL patterns, adding more pages won’t help. Mueller’s framing on cutting page count is the one to internalize: reducing the number of indexable pages without actually improving the site doesn’t make the site better — page-count surgery alone won’t fix a quality-driven Discovered problem. (Paraphrased from his office-hours answer; not verbatim.) Consolidate thin and near-duplicate pages, and make the pages you keep genuinely worth crawling.
Cut crawl waste
Faceted navigationFaceted navigation (faceted search, product filtering) lets visitors refine a list of products or content by attribute — price, color, size, brand, rating. The SEO problem: each filter combination can spawn a distinct crawlable URL, turning a small catalog into millions of near-duplicate pages that waste crawl budget and dilute ranking signals., URL parametersThe `?key=value` data tacked onto the end of a URL after a question mark — used for tracking, sessions, filtering, sorting, and search — and one of the biggest sources of duplicate URLs and wasted crawling in SEO., session IDs, soft 404sA soft 404 is a URL that returns a success status code (usually 200 OK) even though the page is empty, missing, or shows a 'not found' message. It isn't a status code a server sends — it's a label search engines apply after comparing the response code against the rendered content, and they treat the page like a 404 for indexing., and infinite spaces are the classic “Discovered factory” — they spend your crawl capacity on junk URLs so your real content never gets reached. This is where ecommerce and programmatic sites bleed the most. Trimming crawl waste frees capacity and sharpens the quality signal Google reads from your URL patterns. (See crawl budget and spider traps.)
Speed up and stabilize the server
On the capacity side, faster and more stable responses raise your crawl capacity
limit — Google’s own line is that when a site slows down or returns errors, it
crawls less. Fix 5xx errors, cut response times, and remove timeouts.
Earn links to priority pages
External links raise crawl demand for the pages they point at — but slowly. This is a real lever for genuinely important pages, not an instant switch. Don’t expect a backlink to flip a URL out of Discovered overnight.
When (and when not) to use “Request indexing”
Use it for a small number of genuinely important URLs you want crawled sooner. Do not treat it as a fix for thousands of Discovered URLs — it doesn’t scale, and Google explicitly says there’s no need to resubmit. For the sibling Crawled status, Google’s own guidance is that there’s no need to resubmit the URL for crawling; Discovered behaves the same way. Request indexing nudges the queue; it doesn’t change why a page was deprioritized.
When to do nothing
Some Discovered is normal triage — Google found a URL and just hasn’t prioritized it yet, and it may crawl it later on its own. If it’s a handful of genuinely low-value URLs, leaving them is fine. The time to act is when a large or important share of your URLs is stuck in Discovered, because that’s the signal of a fixable capacity-or-demand problem underneath.
When you measure “large or important share,” define the denominator first. The Page Indexing reportThe Google Search Console report (formerly Index Coverage) showing how many of your URLs are indexed vs. not indexed, and grouping the not-indexed ones by reason.’s example list for any status is capped at 1,000 URLs and isn’t guaranteed to show every affected URL — so don’t treat the exported examples as a complete list. Compare the report’s count against your own sitemap/URL inventory (not just the sampled examples) to get an honest share, and prioritize by business importance and traffic potential, not just row count.
Special cases: large, ecommerce, and programmatic sites
This is where Discovered stops being cosmetic. Sites with millions of URLs, faceted navigation, near-duplicate product pages, and infinite parameter spaces generate far more URLs than Google wants to crawl — so a large portion sits in Discovered by design. Here the playbook is crawl-waste reduction first (consolidate, block low value spaces from crawling where appropriate, fix parameter explosions), then internal-linking and quality work to raise demand for the URLs that matter, then server capacity. New sites with weak authority hit a milder version of the same thing: low demand, so weak pages wait.
Where this sits
Discovered is one status in the Page Indexing report, and it’s a crawl-stage problem — which is why the fixes lean on crawl budget, internal linking, and indexing fundamentals. Its sibling, Crawled – currently not indexed, is the quality-stage version of the same frustration. For the upstream stage — how Google discovers and fetches URLs in the first place — see crawling; for the downstream stage, see indexing.
AI summary
A condensed take on the Advanced version:
- What it is: a GSCA free Google service that reports how a site performs in Google Search and surfaces problems with how Google crawls, indexes, and serves it. It's first-party data straight from Google — but you don't need it to appear in results. Page IndexingThe Google Search Console report (formerly Index Coverage) showing how many of your URLs are indexed vs. not indexed, and grouping the not-indexed ones by reason. status meaning Google foundA 302 (\"Found\") is a temporary redirect: it forwards users to a new URL while telling search engines the original URL should stay in the index. It's a weak canonicalization signal, not the zero-equity dead end of SEO folklore. the URL but hasn’t crawled it — the Last Crawl date is empty. It’s a pre-crawl queue state, not a penalty.
- vs Crawled – currently not indexedStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed.: Discovered = not yet fetched (a crawl-priority/capacity signal); Crawled = fetched, and Google is still evaluating it for indexingStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed. (several possible causes, not one “quality verdict”). Getting out of Discovered still doesn’t guarantee indexing.
- Two root causes (Google): crawl capacityThe number of URLs an engine will crawl in a timeframe. — crawlingCrawling is how search engines use automated bots (like Googlebot and Bingbot) to discover URLs and download pages. A page has to be crawlable to be indexed, but crawling on its own isn't a ranking factor. now would overload the server, so Google rescheduled — and crawl demandCrawl demand is the 'want' side of crawl budget — how much a search engine wants to crawl a site or URL, driven by popularity, staleness, and perceived inventory (plus temporary spikes from site moves). It's distinct from crawl rate/capacity, the 'can' side. — your site/pages aren’t worth the crawl effort (quality + internal linkingAn internal link is a hyperlink from one page on a website to another page on the same website. Internal links help search engines discover your pages and pass ranking signals (PageRank and anchor-text context) between them.). Google’s crawl-budget guide targets very large or fast-changing sites; it labels its size thresholds as rough estimates, not exact cutoffs.
- It’s often sitewide, not per-page: Google extrapolates crawl priority from URL patterns and overall site quality — a practitioner inference from pattern behavior, not something the status proves for a single URL.
- Diagnose: Crawl StatsA Google Search Console report (under Settings) that shows how Google has crawled your site over the last 90 days — total requests, download size, and average response time, broken down by response code, file type, Googlebot type, and purpose. It's only available for root-level properties (a Domain property or a URL-prefix property verified at the site's root). + logs for the capacity side; internal-link depth, orphans, and sitewide quality for the demand side. Small/mid-size sites lean demand-bound more often than not — Illyes has been widely quoted (unverified against the original recording) putting it around 90% of sites not needing to worry about crawl budgetThe number of URLs an engine will crawl in a timeframe. — but confirm on your own data rather than assuming from site size.
- Fix: strengthen internal linkingLinks between pages on the same site. and fix orphans; improve quality and consolidate thin/duplicate pages; cut crawl waste (facets, parameters, soft 404sA soft 404 is a URL that returns a success status code (usually 200 OK) even though the page is empty, missing, or shows a 'not found' message. It isn't a status code a server sends — it's a label search engines apply after comparing the response code against the rendered content, and they treat the page like a 404 for indexing., infinite spaces); speed up the server; earn links to priority pages.
- Request indexing nudges a few URLs but doesn’t scale and doesn’t fix the cause — Google says no need to resubmit. Cutting page count without improving quality doesn’t help either.
- Measuring the backlog: the Page Indexing reportThe Google Search Console report (formerly Index Coverage) showing how many of your URLs are indexed vs. not indexed, and grouping the not-indexed ones by reason.’s example rows cap at 1,000 per status and aren’t guaranteed complete — measure share against your own URL inventory, not just the sampled export.
Official documentation
Primary-source documentation from the search engines.
- Page Indexing report — the definitions of “Discovered – currently not indexedStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed.” and “Crawled – currently not indexedA Google Search Console Page Indexing status meaning Googlebot fetched the page but Google decided not to index it — usually a content- or site-quality signal, not a technical error.,” and every other status in the report.
- Optimize your crawl budget — crawl capacityThe number of URLs an engine will crawl in a timeframe. + crawl demandCrawl demand is the 'want' side of crawl budget — how much a search engine wants to crawl a site or URL, driven by popularity, staleness, and perceived inventory (plus temporary spikes from site moves). It's distinct from crawl rate/capacity, the 'can' side.; the guide explicitly names sites with a large share of “Discovered – currently not indexedA Google Search Console Page Indexing status meaning Google knows the URL exists but hasn't crawled it yet — the Last Crawl date is empty. Often a crawl-capacity or crawl-demand (site-quality) signal.” URLs.
- In-Depth Guide to How Google Search Works — crawl → index → serve, URL discoveryURL discovery is how search engines find URLs to crawl — by pull (following links and reading sitemaps) and by push (you notify them via IndexNow, the Indexing API, or WebSub). It's the find step that comes before a page is ever fetched., and why not all pages make it through each stage.
- Crawling and Indexing — the hub for robots, sitemapsA sitemap is a file that lists the pages, images, videos, and other files on your site so search engines can discover them. It helps discovery, but submitting a sitemap doesn't guarantee crawling or indexing., and crawl controls behind crawl-waste fixes.
Bing / Microsoft
- Bing Webmaster Tools — help — Bing doesn’t use the exact “Discovered – currently not indexed” label; it reports indexingStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed. via URL InspectionA Google Search Console feature that reports how Google sees one specific URL on a property you own. By default it shows the last-indexed snapshot; a separate \"Test live URL\" mode fetches the current version. / Site Explorer and governs crawlingCrawling is how search engines use automated bots (like Googlebot and Bingbot) to discover URLs and download pages. A page has to be crawlable to be indexed, but crawling on its own isn't a ranking factor. through crawl quota and content value. Confirm current Bing terminology before citing it.
Quotes from the source
On-the-record statements from Google. Each link is a deep link that jumps to the quoted passage on the source page.
Google — the definition (Page Indexing reportThe Google Search Console report (formerly Index Coverage) showing how many of your URLs are indexed vs. not indexed, and grouping the not-indexed ones by reason.)
- “Discovered - currently not indexedStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed.: The page was foundA 302 (\"Found\") is a temporary redirect: it forwards users to a new URL while telling search engines the original URL should stay in the index. It's a weak canonicalization signal, not the zero-equity dead end of SEO folklore. by Google, but not crawled yet. Typically, Google wanted to crawl the URL but this was expected to overload the site; therefore Google rescheduled the crawl. This is why the last crawl date is empty on the report.” — Google Search ConsoleA free Google service that reports how a site performs in Google Search and surfaces problems with how Google crawls, indexes, and serves it. It's first-party data straight from Google — but you don't need it to appear in results. Help, Page Indexing reportThe Google Search Console report (formerly Index Coverage) showing how many of your URLs are indexed vs. not indexed, and grouping the not-indexed ones by reason.. Jump to quote
Google — the sibling status, for contrast
- “Crawled - currently not indexedA Google Search Console Page Indexing status meaning Googlebot fetched the page but Google decided not to index it — usually a content- or site-quality signal, not a technical error.: The page was crawled by Google but not indexed. It may or may not be indexed in the future; no need to resubmit this URL for crawlingCrawling is how search engines use automated bots (like Googlebot and Bingbot) to discover URLs and download pages. A page has to be crawlable to be indexed, but crawling on its own isn't a ranking factor..” — Google Search ConsoleA free Google service that reports how a site performs in Google Search and surfaces problems with how Google crawls, indexes, and serves it. It's first-party data straight from Google — but you don't need it to appear in results. Help, Page IndexingStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed. report. Jump to quote
Google — crawl capacityThe number of URLs an engine will crawl in a timeframe. (why the server side throttles)
- “Google’s crawlersA crawler — also called a spider or bot — is an automated program that fetches web pages, extracts their links, and queues new URLs to visit. Search engines use crawlers to discover and download content for their index. calculate a crawl capacity limit, which is the maximum number of simultaneous parallel connections that Google can use to crawl a site.” — Google Search Central, Optimize your crawl budgetThe number of URLs an engine will crawl in a timeframe.. Jump to quote
- The same guide applies to “Sites with a large portion of their total URLs classified by Search ConsoleGoogle's free tool for monitoring crawling, indexing, and search performance. as Discovered - currently not indexed.” Jump to quote
”Discovered – currently not indexed” triage checklist
Work top to bottom — confirm what kind of problem you have before you start fixing:
- Confirm the status in GSCA free Google service that reports how a site performs in Google Search and surfaces problems with how Google crawls, indexes, and serves it. It's first-party data straight from Google — but you don't need it to appear in results. URL InspectionA Google Search Console feature that reports how Google sees one specific URL on a property you own. By default it shows the last-indexed snapshot; a separate \"Test live URL\" mode fetches the current version. — is the Last Crawl date empty? (Empty = genuinely Discovered, not Crawled.)
- Capacity check: review Crawl StatsA Google Search Console report (under Settings) that shows how Google has crawled your site over the last 90 days — total requests, download size, and average response time, broken down by response code, file type, Googlebot type, and purpose. It's only available for root-level properties (a Domain property or a URL-prefix property verified at the site's root). (average response time, host
status) and server logsLog file analysis is reading a web server's raw access logs to see exactly which URLs search engine crawlers actually requested, when, how often, and what status code they got. Unlike crawl tools or Search Console, logs are the unsampled, ground-truth record of what really happened. for slow responses, timeouts, and
5xxspikes. - Demand check: are the affected URLs orphaned or buried deep in the architecture? Map their internal-link depth.
- Quality check: are they thin, near-duplicateThe same or very similar primary content reachable at more than one URL. There's no general duplicate content penalty — the real costs are possible signal dilution, the wrong URL getting chosen, and less-efficient crawling., or auto-generated URL patterns rather than genuinely distinct pages?
- Crawl-waste check: is faceted nav, parameters, session IDs, soft 404sA soft 404 is a URL that returns a success status code (usually 200 OK) even though the page is empty, missing, or shows a 'not found' message. It isn't a status code a server sends — it's a label search engines apply after comparing the response code against the rendered content, and they treat the page like a 404 for indexing., or an infinite space inflating your URL count?
- Internal linksAn internal link is a hyperlink from one page on a website to another page on the same website. Internal links help search engines discover your pages and pass ranking signals (PageRank and anchor-text context) between them.: important Discovered pages are linked from pages Google already crawls (homepage, hubs, nav) and aren’t many clicks deep.
- SitemapA sitemap is a file that lists the pages, images, videos, and other files on your site so search engines can discover them. It helps discovery, but submitting a sitemap doesn't guarantee crawling or indexing.: affected URLs are in your XML sitemapAn XML sitemap is a UTF-8 file listing the canonical URLs on your site (with optional lastmod) so search engines can discover and prioritize them. It's a discovery and diagnostic aid, not a guarantee of indexing — and Google ignores its priority and changefreq tags. (helps discovery — not a priority lever).
- Server health: fast, stable responses;
5xx/timeouts minimized. - Decide scope: a handful of low-value URLs → fine to leave. A large or important share stuck → fix the demand/capacity cause.
- Request indexingStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed. only for a few genuinely important URLs — not as a bulk fix.
The mental models
1. Discovered = foundA 302 (\"Found\") is a temporary redirect: it forwards users to a new URL while telling search engines the original URL should stay in the index. It's a weak canonicalization signal, not the zero-equity dead end of SEO folklore., not fetched. The empty Last Crawl date is the whole diagnosis. If something was fetched, it isn’t Discovered — it’s Crawled. Get this right first; everything downstream depends on it.
2. Capacity vs demand. There are only two reasons a URL sits in Discovered: Google couldn’t crawl it (server capacity) or Google didn’t want to crawl it (demand — quality, linking, priority). Diagnose which one you have before you touch anything. Crawl StatsA Google Search Console report (under Settings) that shows how Google has crawled your site over the last 90 days — total requests, download size, and average response time, broken down by response code, file type, Googlebot type, and purpose. It's only available for root-level properties (a Domain property or a URL-prefix property verified at the site's root). and logs tell you capacity; link depth, orphans, and sitewide quality tell you demand.
3. It’s often a sitewide signal, not a per-page bug. Google extrapolates crawl priority from URL patterns and overall quality — that’s a practitioner inference from pattern behavior, not a fact the status proves for one URL. Discovered is more often a signal about a category of your pages than a flaw in one page. Fix the pattern, not just the page — but confirm the pattern with your own data first.
4. CrawlingCrawling is how search engines use automated bots (like Googlebot and Bingbot) to discover URLs and download pages. A page has to be crawlable to be indexed, but crawling on its own isn't a ranking factor. is a gate, not a guarantee. Getting a page out of Discovered only earns it a fetch. It can still land in Crawled – currently not indexedStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed. and never get indexed. Plan for the quality stage, not just the crawl stage.
5. Small and mid-size sites are more often demand-bound. True crawl-capacity ceilings mostly bite very large sites. If you’re small or mid-size and seeing Discovered, demand — internal linkingAn internal link is a hyperlink from one page on a website to another page on the same website. Internal links help search engines discover your pages and pass ranking signals (PageRank and anchor-text context) between them. and quality — is the more likely first check, not a server wall. Confirm with Crawl Stats and logs rather than assuming it from site size alone.
6. Request indexing nudges; it doesn’t cure. It moves a few URLs up the queue. It changes nothing about why they were deprioritized, and it doesn’t scale. Use the structural levers for the real fix.
Discovered vs Crawled — and the fix-it map
Discovered vs Crawled – currently not indexedStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed.
| Discovered – currently not indexedA Google Search Console Page Indexing status meaning Google knows the URL exists but hasn't crawled it yet — the Last Crawl date is empty. Often a crawl-capacity or crawl-demand (site-quality) signal. | Crawled – currently not indexedA Google Search Console Page Indexing status meaning Googlebot fetched the page but Google decided not to index it — usually a content- or site-quality signal, not a technical error. | |
|---|---|---|
| What happened | FoundA 302 (\"Found\") is a temporary redirect: it forwards users to a new URL while telling search engines the original URL should stay in the index. It's a weak canonicalization signal, not the zero-equity dead end of SEO folklore., not yet fetched | Fetched, not kept |
| Last Crawl date | Empty | Present |
| Stage | Pre-crawl (queue) | Post-crawl (index decision) |
| Primary signal | Crawl priority / capacity | IndexingStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed. evaluation (several possible causes) |
| Typical cause | Weak internal linksAn internal link is a hyperlink from one page on a website to another page on the same website. Internal links help search engines discover your pages and pass ranking signals (PageRank and anchor-text context) between them., crawl waste, server load, low demand | Duplication, thin contentThin content is web content that provides little or no value to users. Google's spam policies name it 'thin content with little or no added value' — and it's about value per page, not word count., canonicalizationHow search engines pick one canonical URL among duplicates and consolidate signals onto it. to another URL, and more |
| First lever | Internal linkingLinks between pages on the same site., cut crawl waste, server speed | Improve/consolidate the page itself |
| Resubmit needed? | No (nudge a few priority URLs only) | No |
Capacity vs demand — which problem is it?
| Symptom | Likely cause | First fix |
|---|---|---|
Slow Crawl StatsA Google Search Console report (under Settings) that shows how Google has crawled your site over the last 90 days — total requests, download size, and average response time, broken down by response code, file type, Googlebot type, and purpose. It's only available for root-level properties (a Domain property or a URL-prefix property verified at the site's root)., 5xx/timeout spikes in logs | Crawl capacity | Speed up / stabilize the server |
| URLs orphaned or buried deep | Crawl demand | Internal linking, pull pages shallower |
| Thin / near-duplicateThe same or very similar primary content reachable at more than one URL. There's no general duplicate content penalty — the real costs are possible signal dilution, the wrong URL getting chosen, and less-efficient crawling. / auto-generated patterns | Crawl demand (quality) | Consolidate, improve, prune |
| Millions of facet/parameter URLs | Crawl waste | Reduce infinite spaces, manage parameters |
| Small site, few Discovered URLs | Normal triage | Often fine to leave |
Fast facts
- Discovered means empty Last Crawl date — the one fact that distinguishes it.
- Google’s two drivers: crawl capacityThe number of URLs an engine will crawl in a timeframe. + crawl demandCrawl demand is the 'want' side of crawl budget — how much a search engine wants to crawl a site or URL, driven by popularity, staleness, and perceived inventory (plus temporary spikes from site moves). It's distinct from crawl rate/capacity, the 'can' side..
- Gary Illyes has been widely quoted (industry coverage, not independently verified) as saying roughly 90% of sites don’t need to worry about crawl budgetThe number of URLs an engine will crawl in a timeframe. — a directional signal, not a confirmed stat. On small/mid-size sites, persistent Discovered is more often a quality/linking problem than a capacity ceiling.
- Request indexing doesn’t scale and doesn’t fix the cause; no need to resubmit.
- Getting crawled does not guarantee indexing.
- The Page Indexing reportThe Google Search Console report (formerly Index Coverage) showing how many of your URLs are indexed vs. not indexed, and grouping the not-indexed ones by reason.’s example rows cap at 1,000 per status and aren’t guaranteed complete — measure share against your own URL inventory, not the sample.
Why hasn’t Google crawled these URLs?
Discovered – currently not indexed diagnosis
Patrick's relevant free tools
- SEO Incident Simulator — Practice thirty deterministic technical SEO incident investigations — indexability, crawl controls, redirects, sitemaps, markup, caching, DNS, bot verification, rendering, hreflang, and faceted navigation — with clearly labeled fixture evidence and Find → Fix → Verify handoffs.
- Google Index Checker — Check one URL’s observable indexability blockers, or reconcile sitemap, crawl, and supplied Search Console evidence across a URL set before verifying Google’s actual state in URL Inspection.
Tools for separating capacity from demand
- Log File Analyzer — check whether GooglebotGooglebot is Google's web crawler — the software that fetches pages so Google can index and rank them. It comes in two variants, Googlebot Smartphone (primary, under mobile-first indexing) and Googlebot Desktop, and runs an evergreen Chromium renderer. requests are falling, which URL patterns consume requests, and whether errors are present.
- Robots.txt Tester — rule out an access block before treating a missing crawl as a scheduling decision.
- Sitemap Validator — verify that priority canonical URLsHow search engines pick one canonical URL among duplicates and consolidate signals onto it. are present and the sitemapA sitemap is a file that lists the pages, images, videos, and other files on your site so search engines can discover them. It helps discovery, but submitting a sitemap doesn't guarantee crawling or indexing. is fetchable and structurally sound.
- Link Analyzer — inspect whether the affected page has a crawlable internal path instead of existing only in a sitemapA sitemap is a file that lists the pages, images, videos, and other files on your site so search engines can discover them. It helps discovery, but submitting a sitemap doesn't guarantee crawling or indexing..
Validation tests
Test: internal-link and sitemap fix produces a crawl
Test to run — publish the internal linkAn internal link is a hyperlink from one page on a website to another page on the same website. Internal links help search engines discover your pages and pass ranking signals (PageRank and anchor-text context) between them. and sitemapA sitemap is a file that lists the pages, images, videos, and other files on your site so search engines can discover them. It helps discovery, but submitting a sitemap doesn't guarantee crawling or indexing. correction, then check the URL’s server-log history and URL InspectionA Google Search Console feature that reports how Google sees one specific URL on a property you own. By default it shows the last-indexed snapshot; a separate \"Test live URL\" mode fetches the current version. status. Expected result — GooglebotGooglebot is Google's web crawler — the software that fetches pages so Google can index and rank them. It comes in two variants, Googlebot Smartphone (primary, under mobile-first indexing) and Googlebot Desktop, and runs an evergreen Chromium renderer. requests the canonical URLHow search engines pick one canonical URL among duplicates and consolidate signals onto it. and the Last crawl field is no longer empty. Failure interpretation — the URL still has weak discovery/priority signals, a conflicting variant, or is part of a broader crawl-demand problem. Monitoring window — use the site’s normal crawl cadence; compare against similar priority pages rather than assuming an instant visit. Rollback trigger — remove the new link only if it creates an unintended navigation or duplicate-URL path; otherwise diagnose the remaining signals instead of undoing discovery.
Test: server-capacity repair restores crawling
Test to run — deploy the server fix and compare GooglebotGooglebot is Google's web crawler — the software that fetches pages so Google can index and rank them. It comes in two variants, Googlebot Smartphone (primary, under mobile-first indexing) and Googlebot Desktop, and runs an evergreen Chromium renderer. request volume, response time, and error responses in server logsLog file analysis is reading a web server's raw access logs to see exactly which URLs search engine crawlers actually requested, when, how often, and what status code they got. Unlike crawl tools or Search Console, logs are the unsampled, ground-truth record of what really happened.. Expected result — successful crawlerA crawler — also called a spider or bot — is an automated program that fetches web pages, extracts their links, and queues new URLs to visit. Search engines use crawlers to discover and download content for their index. requests recover without the prior error or latency pattern. Failure interpretation — the capacity constraint remains, or crawl demandCrawl demand is the 'want' side of crawl budget — how much a search engine wants to crawl a site or URL, driven by popularity, staleness, and perceived inventory (plus temporary spikes from site moves). It's distinct from crawl rate/capacity, the 'can' side. rather than capacity is the limiting factor. Monitoring window — compare multiple crawl cycles and the same weekday/time pattern used for the pre-fix baseline. Rollback trigger — revert if the deployment increases crawlerA crawler — also called a spider or bot — is an automated program that fetches web pages, extracts their links, and queues new URLs to visit. Search engines use crawlers to discover and download content for their index.-facing errors or response latency.
How to measure the problem
Discovered-not-indexed population
Metric — count and share of submitted canonical URLsHow search engines pick one canonical URL among duplicates and consolidate signals onto it. in this status, measured against your own sitemapA sitemap is a file that lists the pages, images, videos, and other files on your site so search engines can discover them. It helps discovery, but submitting a sitemap doesn't guarantee crawling or indexing./URL inventory as the denominator — not just the report’s sampled example rows. What it tells you — whether the backlog is growing faster than Google crawls it. How to pull it — export the Page Indexing reportThe Google Search Console report (formerly Index Coverage) showing how many of your URLs are indexed vs. not indexed, and grouping the not-indexed ones by reason. and segment by sitemapA sitemap is a file that lists the pages, images, videos, and other files on your site so search engines can discover them. It helps discovery, but submitting a sitemap doesn't guarantee crawling or indexing. or template. The report’s example list caps at 1,000 URLs per status and isn’t guaranteed to be exhaustive, so treat exported examples as a sample, not a census. Benchmark / realistic range — establish a baseline by template; the useful target is a shrinking backlog for priority inventory, not a universal percentage. Cadence — weekly during remediation, then monthly.
Time from discovery to first crawl
Metric — elapsed time between publication/sitemap inclusion and the first GooglebotGooglebot is Google's web crawler — the software that fetches pages so Google can index and rank them. It comes in two variants, Googlebot Smartphone (primary, under mobile-first indexing) and Googlebot Desktop, and runs an evergreen Chromium renderer. request. What it tells you — whether priority and capacity changes improve scheduling. How to pull it — join publishing or sitemap timestamps to server-log first-seen requests. Benchmark / realistic range — compare like-for-like page types on your own site; crawl cadence varies too much for a universal threshold. Cadence — review monthly or after a material template/server change.
Crawler success and waste
Metric — successful GooglebotGooglebot is Google's web crawler — the software that fetches pages so Google can index and rank them. It comes in two variants, Googlebot Smartphone (primary, under mobile-first indexing) and Googlebot Desktop, and runs an evergreen Chromium renderer. requests to important URLs versus errors and low-value URL patterns. What it tells you — whether crawl capacityThe number of URLs an engine will crawl in a timeframe. is being spent on the inventory you care about. How to pull it — segment server logsLog file analysis is reading a web server's raw access logs to see exactly which URLs search engine crawlers actually requested, when, how often, and what status code they got. Unlike crawl tools or Search Console, logs are the unsampled, ground-truth record of what really happened. by status code and URL pattern. Benchmark / realistic range — use the pre-change mix as the baseline and require improvement in the priority share without increasing errors. Cadence — weekly while diagnosing; monthly once stable.
Prompts for backlog analysis
Find template-level causes
Group this export of “Discovered – currently not indexedStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed.” URLs by template and URL pattern. For each group, compare sitemapA sitemap is a file that lists the pages, images, videos, and other files on your site so search engines can discover them. It helps discovery, but submitting a sitemap doesn't guarantee crawling or indexing. presence, internal-link count, publication date, and server-log first-seen data. Rank capacity, crawl-waste, and crawl-demand hypotheses by evidence. Do not claim a cause when the required evidence is absent.
Prioritize a remediation sample
Select a representative test set from these affected URLs: high-priority pages, low-priority pages, recent pages, old pages, and each major template. Propose one change per hypothesis and state the observable pass signal and rollback trigger. Data: [paste rows].
Test yourself
Resources worth your time
My related writing
- How to Fix “Discovered - currently not indexed” — the Ahrefs guide I review: the five diagnostic areas (crawl budgetThe number of URLs an engine will crawl in a timeframe., content quality, internal linkingAn internal link is a hyperlink from one page on a website to another page on the same website. Internal links help search engines discover your pages and pass ranking signals (PageRank and anchor-text context) between them., backlinks, technical problems) and the fixes for each.
- The Beginner’s Guide to Technical SEO — where crawlingCrawling is how search engines use automated bots (like Googlebot and Bingbot) to discover URLs and download pages. A page has to be crawlable to be indexed, but crawling on its own isn't a ranking factor. and indexingStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed. fit in the bigger picture.
Official
- Page Indexing report (Google) — the source for both status definitions.
- Optimize your crawl budget (Google) — crawl capacity + demand, and the doc that explicitly names this status.
From others
- Google On Fixing “Discovered Currently Not Indexed” (Search Engine Journal) — coverage of Mueller’s capacity-vs-quality framing.
- Understanding and resolving “Discovered - currently not indexed” (Search Engine Land) — Dan Taylor’s diagnostic walkthrough.
- How To Fix “Discovered – Currently Not Indexed” (Onely) — technical deep-dive on diagnosing crawl-capacity vs. quality causes, with server-log analysis.
- Google On Discovered – Currently Not Indexed (Search Engine Roundtable) — Barry Schwartz covering Google rep commentary on this status.
- r/TechSEO — the community for crawl/index debugging.
Discovered – currently not indexed
A Google Search Console Page Indexing status meaning Google knows the URL exists but hasn't crawled it yet — the Last Crawl date is empty. Often a crawl-capacity or crawl-demand (site-quality) signal.
Related: Crawled – currently not indexed, Crawl Budget, Indexing
Discovered – currently not indexed
“Discovered – currently not indexedStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed.” is a status in Google Search ConsoleA free Google service that reports how a site performs in Google Search and surfaces problems with how Google crawls, indexes, and serves it. It's first-party data straight from Google — but you don't need it to appear in results.’s Page Indexing reportThe Google Search Console report (formerly Index Coverage) showing how many of your URLs are indexed vs. not indexed, and grouping the not-indexed ones by reason.. It means Google has foundA 302 (\"Found\") is a temporary redirect: it forwards users to a new URL while telling search engines the original URL should stay in the index. It's a weak canonicalization signal, not the zero-equity dead end of SEO folklore. the URL — usually from a sitemapA sitemap is a file that lists the pages, images, videos, and other files on your site so search engines can discover them. It helps discovery, but submitting a sitemap doesn't guarantee crawling or indexing. or an internal or external link — but hasn’t crawled it yet. The giveaway is that the Last Crawl date is empty: nothing has been fetched.
Google’s own explanation is that it wanted to crawl the URL, but doing so was expected to overload the site, so it rescheduled the crawl. It can also signal weak crawl demandCrawl demand is the 'want' side of crawl budget — how much a search engine wants to crawl a site or URL, driven by popularity, staleness, and perceived inventory (plus temporary spikes from site moves). It's distinct from crawl rate/capacity, the 'can' side. — Google’s systems not prioritizing the page, which practitioners commonly trace back to overall site quality and internal linkingAn internal link is a hyperlink from one page on a website to another page on the same website. Internal links help search engines discover your pages and pass ranking signals (PageRank and anchor-text context) between them., though that’s an inference from URL-pattern behavior rather than something the status proves on its own.
The one fact that separates it from its sibling status, “Crawled – currently not indexedA Google Search Console Page Indexing status meaning Googlebot fetched the page but Google decided not to index it — usually a content- or site-quality signal, not a technical error.,” is the crawl itself: Discovered means not yet fetched (empty Last Crawl date), while Crawled means Google fetched the page and is evaluating it for indexingStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed. (several possible causes, not a single quality verdict). One is a crawl-priority problem; the other is an indexing-evaluation problem.
On small and mid-size sites, persistent Discovered is more often a demand issue than a true crawl-capacity ceiling — it tends to point at quality, internal linkingLinks between pages on the same site., or crawl waste, though confirm that against your own Crawl StatsA Google Search Console report (under Settings) that shows how Google has crawled your site over the last 90 days — total requests, download size, and average response time, broken down by response code, file type, Googlebot type, and purpose. It's only available for root-level properties (a Domain property or a URL-prefix property verified at the site's root). and logs rather than assuming it. “Request indexing” can nudge a handful of priority URLs but doesn’t scale to thousands and doesn’t fix the underlying cause.
Related: Crawled – currently not indexed, Crawl Budget, Indexing
Build-time retrieval analysis plus live signals for this exact article. The automatic chunk report includes a deterministic readiness score and is ready without a model download.
Search Console
sampleGA4 traffic (28d)
sampleCloudflare traffic (7d)
sampledCrUX field data (28d, phone)
sampleGoogle NLP entities
localChangelog
Updated Jul 17, 2026.
Editorial summary and recorded change details.Summary
Tightened several unverified frequency claims (the ~90% crawl-budget stat, 'usually sitewide' / 'almost always demand' framing, and URL-pattern-quality inference) into clearly hedged, evidence-first language, added the Page Indexing report's 1,000-example sampling caveat, and reframed 'Crawled – currently not indexed' as a multi-cause indexing evaluation rather than a single quality verdict.
Change details
-
Added a note that Google's crawl-budget guide's site-size audience thresholds are explicitly rough estimates, not exact cutoffs.
-
Added the Page Indexing report's 1,000-example-per-status sampling cap and an inventory-denominator requirement to the backlog-measurement guidance.
Full comparison unavailable — no prior snapshot was archived for this revision.