Orphan Pages

Orphan pages have no internal links, so search engines struggle to find them and they get zero internal PageRank. How to spot and fix them.

First published: Jun 26, 2026 · Last updated: Jul 18, 2026 · Advanced
demand #3 in Internal Links#12 in Website Structure#132 in Technical SEO#181 on the site
1 evidence signal on this page

An orphan page is one nothing on your site links to internally. Because Google discovers pages mainly by following links, an orphan with no sitemap entry and no backlinks is effectively invisible — and even when it's found via a sitemap, it gets zero internal PageRank and often lands in 'Discovered — currently not indexed.' A crawler alone can't find orphans (it only follows links it can reach); you have to compare the crawl against your sitemap, analytics, and backlink data. Fix the ones you want ranked by adding contextual internal links; noindex the intentional ones (PPC and thank-you pages); redirect or delete the rest.

TL;DR — Google discoversGoogle Discover is a personalized, mobile-first content feed built into the Google app, Chrome's mobile New Tab page, and google.com that surfaces articles and videos based on a user's interests and activity — not a response to a search query. There's nothing to 'rank' for in the traditional sense; eligibility is governed by Discover's content policies plus the same helpful-content, image, and page-experience signals Google Search already uses. URLs three ways — following links (primary), sitemapsA sitemap is a file that lists the pages, images, videos, and other files on your site so search engines can discover them. It helps discovery, but submitting a sitemap doesn't guarantee crawling or indexing. (supplemental), and re-crawlingCrawling is how search engines use automated bots (like Googlebot and Bingbot) to discover URLs and download pages. A page has to be crawlable to be indexed, but crawling on its own isn't a ranking factor. pages it already knows. An orphan page (no internal linksAn internal link is a hyperlink from one page on a website to another page on the same website. Internal links help search engines discover your pages and pass ranking signals (PageRank and anchor-text context) between them.) with no sitemapA sitemap is a file that lists the pages, images, videos, and other files on your site so search engines can discover them. It helps discovery, but submitting a sitemap doesn't guarantee crawling or indexing. entry and no backlinks has no discovery path at all. Even when a sitemap gets it crawled, it receives zero internal PageRankPageRank is Google's original recursive link-graph algorithm: a page's score depends on the scores of the pages linking to it, and in the published model each page's score is split across its outbound links (the simplified version: links are weighted votes). Google says it's evolved since launch but still part of its core ranking systems., which is why orphans cluster in “Discovered — currently not indexedStoring a crawled page in the search index so it can appear in results. Crawled is not the same as indexed — Google selects what to keep, and indexing isn't guaranteed..” A crawl alone can’t find orphans — you compare the crawl against a superset of known URLs (sitemap, analytics, GSCA free Google service that reports how a site performs in Google Search and surfaces problems with how Google crawls, indexes, and serves it. It's first-party data straight from Google — but you don't need it to appear in results., backlink index). Fix the ones you want ranked with contextual internal linksAn internal link is a hyperlink from one page on a website to another page on the same website. Internal links help search engines discover your pages and pass ranking signals (PageRank and anchor-text context) between them.; noindex the intentional ones; redirectA redirect sends browsers and crawlers from a requested URL to a different one. An HTTP redirect specifically is a 3xx status code paired with a Location header; meta refresh and JavaScript redirects achieve a similar navigation without being a 3xx response themselves. Permanent redirects (301/308) are Google's signal the target should be canonical; temporary ones (302/303/307) aren't. or delete the rest. In our 1M-domain study, 66.2% of sites had pages with only one dofollow internal link — orphans are the extreme end of a very common internal-link-poverty problem.

Evidence for this claim Google discovers pages through links and recommends that every page you care about have a link from at least one other page on the site. Scope: Current Google link and discovery guidance. Confidence: high · Verified: Google Search Central: Link best practices Evidence for this claim Sitemaps can help search engines discover URLs but do not guarantee crawling or indexing and do not replace navigational links. Scope: Current Google sitemap behavior. Confidence: high · Verified: Google Search Central: Learn about sitemaps

What makes a page an orphan

An orphan page is one that no other page on your site linksSitelinks are extra links from the same domain that Google clusters together under a single search result, usually for branded or navigational queries. They're generated entirely algorithmically — there's no way to add, edit, or guarantee them. to internally. That’s the whole definition — it’s about incoming internal links, nothing else. A page can have a pile of external backlinks and still be an orphan internally. A page can sit in your XML sitemapAn XML sitemap is a UTF-8 file listing the canonical URLs on your site (with optional lastmod) so search engines can discover and prioritize them. It's a discovery and diagnostic aid, not a guarantee of indexing — and Google ignores its priority and changefreq tags. and still be an orphan. Orphan status is measured purely against your site’s internal link graph.

It helps to think of internal linkingLinks between pages on the same site. as a spectrum rather than a binary. Truly orphaned (zero internal links) is the worst case; one internal link is barely better; a well-linked page sits at the healthy end. When we studied over a million domains for the Ahrefs Site Audit study, 66.2% of sites had pages with only a single dofollow incoming internal link. As I put it then: “At least it’s not a completely orphaned page. But if it’s a page that you want to rank, you may want to add some more internal links.” Orphans are just the extreme end of that same under-linking problem.

Why orphan pages are an SEO problem

Search engines can’t reliably discover them

Google is explicit about how it finds URLs. “There isn’t a central registry of all web pages, so Google must constantly look for new and updated pages.” It does that three ways: pages it already knows from prior crawls, “when Google extracts a link from a known page to a new page,” and “when you submit a list of pages (a sitemap) for Google to crawl.”

Link-following is the primary mechanism; sitemaps are supplemental, not a substitute. Google’s own link best-practices guidance is blunt: “Every page you care about should have a link from at least one other page on your site.” An orphan with no sitemap entry and no external backlinks has none of the three discovery paths — it is genuinely invisible to GooglebotGooglebot is Google's web crawler — the software that fetches pages so Google can index and rank them. It comes in two variants, Googlebot Smartphone (primary, under mobile-first indexing) and Googlebot Desktop, and runs an evergreen Chromium renderer..

John Mueller has called internal linking “super critical for SEO… one of the biggest things that you can do on a website to kind of guide Google and guide visitors to the pages that you think are important.” His framing is about link distance from your important hub pages, not just raw link count — an orphan has infinite link distance, because it isn’t reachable from your link graph at all.

They receive no internal PageRank

PageRankPageRank is Google's original recursive link-graph algorithm: a page's score depends on the scores of the pages linking to it, and in the published model each page's score is split across its outbound links (the simplified version: links are weighted votes). Google says it's evolved since launch but still part of its core ranking systems. still underpins Google’s link-based authority, and internal links are the plumbing that moves it around your site. An orphan page is disconnected from that plumbing, so it receives zero internal PageRank no matter how good its content is.

This is the angle most people miss, and it cuts both ways. A page with strong external backlinks but no internal links isn’t just under-powered — it’s a PageRank leak. Authority flows into that page from the outside and then has nowhere to go, because no internal links carry it onward to the rest of your site. Add internal links and you both strengthen the orphan and let its inbound equity circulate.

They cluster in “Discovered — currently not indexed”

When Google knows a URL exists (from a sitemap or an external link) but decides not to index it, GSC reports “Discovered — currently not indexed.” Orphan pages are a classic cause: Google found the URL, but with no internal links pointing at it, there’s no signal that the page matters, so it gets deprioritized. Adding internal links from relevant pages is one of the most reliable ways to nudge those URLs out of that bucket. (It’s not the only cause — thin contentThin content is web content that provides little or no value to users. Google's spam policies name it 'thin content with little or no added value' — and it's about value per page, not word count. and crawl budget play in too — but it’s one of the first things I check.)

They can waste crawl budget — but mostly on large sites

For the average site this barely registers. On very large sites (think 100k+ pages, or large numbers of rapidly changing URLs), big pools of orphaned and near-orphaned URLs consume crawl budgetThe number of URLs an engine will crawl in a timeframe. without earning their keep. Most sites genuinely don’t need to think about this; if you’re below that scale, treat orphans as a discovery and PageRank problem, not a crawl-budget one.

Common causes

  • Migrations and redesigns — URLs survive the move but the links that pointed to them don’t get rebuilt.
  • Navigation changes — a page gets pulled from the menu and nothing replaces the link.
  • CMSA content management system (CMS) is software that lets users create, manage, and publish digital content — like blog posts and pages — without writing raw code. WordPress, Drupal, and Joomla are the most common open-source CMS platforms. workflow gaps — a post is published without being assigned to a category or tag, so the only link path (the category archive) never gets created. WordPress and similar CMSes generate a lot of paginationPagination splits a large set of content — product listings, blog archives, search results — across multiple sequentially numbered URLs. For SEO, each paginated page should be crawlable, indexable, and self-canonical; Google no longer uses rel=prev/next, but Bing still does. and archive URLs that can strand content this way.
  • Discontinued productsA discontinued product is an item you'll never sell again — the manufacturer stopped making it, or you dropped the line. The SEO decision is end-of-life: 301-redirect the URL to a genuinely similar replacement or the closest relevant category if it earned links or traffic, 404/410 it if it didn't, or keep it live as a Discontinued tombstone page only when it still helps users. This is distinct from a temporary out-of-stock product, which you keep live at 200. — the product page is left live after it’s removed from listings.
  • Campaign and PPC landing pages — built to be reached from an ad, never linked from the site (often intentional).
  • Staging, test, and siloed microsite content — left publicly accessible and never wired into the main link graph.

How to find orphan pages

The core method is a two-list comparison, because no crawl can find orphans on its own:

  1. List 1 — reachable pages. Run a site crawlerA crawler — also called a spider or bot — is an automated program that fetches web pages, extracts their links, and queues new URLs to visit. Search engines use crawlers to discover and download content for their index. that starts at your homepage and follows internal links. This is everything Google could discover by link-following.
  2. List 2 — all known pages. Gather every URL you know exists: your XML sitemap, Google Analytics (pages with sessions), GSC pages, your backlink index, and server logs.
  3. Orphans = in List 2 but not List 1.

In Ahrefs Site Audit, this is built in — it seeds the crawl from your sitemaps, the Ahrefs backlink index, and connected GA/GSC data, then flags any URL it knows about but couldn’t reach via internal links as an “Orphan page.” In Screaming Frog, you upload your sitemap, GA, and GSC data as URL sources, run the crawl in list/crawl mode, then use the crawl analysis to surface URLs that appear in those sources but weren’t found by link-following. With GSC alone you can’t see site structure, but URLs sitting in “Discovered — currently not indexed” that don’t appear in your crawl are strong orphan suspects.

TIP Build internal-link candidates for the orphan pages you keep

Supply the orphan candidates alongside reachable cluster pages. A similarity score can prioritize possible sources, but only an editor can confirm that a link belongs in the passage.

Create a review queue with my free Internal Link Cluster Visualizer Free

  1. First prove orphan status by comparing the crawl with sitemaps, analytics, GSC, or backlink data.
  2. Supply valuable orphan candidates and relevant reachable pages to generate possible connections.
  3. Add only contextually useful links, then recrawl to confirm each kept page is reachable.

How to fix them

Triage by value first — not every orphan deserves a link.

Before you triage, confirm the candidate itself is worth linking to: check its response status, its canonical target, and whether it’s set to noindex. Linking to a URL that redirects, 404s, canonicalizes elsewhere, or is deliberately excluded from the index isn’t a fix — it just adds a broken or wasted link. Only pages that resolve cleanly and are meant to be indexed belong in the “add links” bucket below.

  • Valuable pages you want to rank → add internal links. Find topically relevant pages and link from them with descriptive anchor textAnchor text is the visible, clickable text of a hyperlink. It tells readers what they'll find on the other end and gives search engines context about the linked page.. Prioritize linking from pages that already have internal authority (hub and cluster pages), and make sure the page is in your XML sitemapAn XML sitemap is a UTF-8 file listing the canonical URLs on your site (with optional lastmod) so search engines can discover and prioritize them. It's a discovery and diagnostic aid, not a guarantee of indexing — and Google ignores its priority and changefreq tags. too.
  • Intentional orphans (PPC, thank-you, some legal pages) → noindex. These don’t need internal links; they need to be kept out of the index cleanly. Adding links would be the wrong fix. Give each category of intentional orphan a named owner and a documented reason it’s excluded, so the next audit doesn’t re-flag it as an accidental one.
  • Low-value pages → redirect or delete. 301 to the most relevant alternative, or remove it entirely if there’s nothing relevant and no backlinks worth preserving.

Prevent recurrence: bake an internal-linking step into your publishing workflow, assign categories/tags before publishing, and re-audit after every migration, redesign, or navigation change. This is the same internal-link discipline that governs crawl depthCrawl depth usually means click depth — how many clicks it takes to reach a page from the homepage by following internal links. It can also mean a crawler setting that limits how many levels deep a crawl goes before it stops. and how PageRank flows through your architecture — orphan prevention is just the floor of good internal linking.

A note on the related-but-different term: an orphan page has no incoming internal links; a dead-end page has no outgoing internal links. They’re separate issues, and a page can be both.

Add an expert note

Pin an expert quote

New person? Create their unclaimed profile at /admin/experts/ → Pin a quote first.