Orphan Pages

Orphan pages have no internal links, so search engines struggle to find them and they get zero internal PageRank. How to spot and fix them.

First published: Jun 26, 2026 · Last updated: Jul 18, 2026 · Advanced
demand #2 in Internal Links#10 in Website Structure#145 in Technical SEO#189 on the site
1 evidence signal on this page

An orphan page is one nothing on your site links to internally. Because Google discovers pages mainly by following links, an orphan with no sitemap entry and no backlinks is effectively invisible — and even when it's found via a sitemap, it gets zero internal PageRank and often lands in 'Discovered — currently not indexed.' A crawler alone can't find orphans (it only follows links it can reach); you have to compare the crawl against your sitemap, analytics, and backlink data. Fix the ones you want ranked by adding contextual internal links; noindex the intentional ones (PPC and thank-you pages); redirect or delete the rest.

TL;DR — Google discovers URLs three ways — following links (primary), sitemaps (supplemental), and re-crawling pages it already knows. An orphan page (no internal links) with no sitemap entry and no backlinks has no discovery path at all. Even when a sitemap gets it crawled, it receives zero internal PageRank, which is why orphans cluster in “Discovered — currently not indexed.” A crawl alone can’t find orphans — you compare the crawl against a superset of known URLs (sitemap, analytics, GSC, backlink index). Fix the ones you want ranked with contextual internal links; noindex the intentional ones; redirect or delete the rest. In our 1M-domain study, 66.2% of sites had pages with only one dofollow internal link — orphans are the extreme end of a very common internal-link-poverty problem.

Evidence for this claim Google discovers pages through links and recommends that every page you care about have a link from at least one other page on the site. Scope: Current Google link and discovery guidance. Confidence: high · Verified: Google Search Central: Link best practices Evidence for this claim Sitemaps can help search engines discover URLs but do not guarantee crawling or indexing and do not replace navigational links. Scope: Current Google sitemap behavior. Confidence: high · Verified: Google Search Central: Learn about sitemaps

What makes a page an orphan

An orphan page is one that no other page on your site links to internally. That’s the whole definition — it’s about incoming internal links, nothing else. A page can have a pile of external backlinks and still be an orphan internally. A page can sit in your XML sitemap and still be an orphan. Orphan status is measured purely against your site’s internal link graph.

It helps to think of internal linking as a spectrum rather than a binary. Truly orphaned (zero internal links) is the worst case; one internal link is barely better; a well-linked page sits at the healthy end. When we studied over a million domains for the Ahrefs Site Audit study, 66.2% of sites had pages with only a single dofollow incoming internal link. As I put it then: “At least it’s not a completely orphaned page. But if it’s a page that you want to rank, you may want to add some more internal links.” Orphans are just the extreme end of that same under-linking problem.

Why orphan pages are an SEO problem

Search engines can’t reliably discover them

Google is explicit about how it finds URLs. “There isn’t a central registry of all web pages, so Google must constantly look for new and updated pages.” It does that three ways: pages it already knows from prior crawls, “when Google extracts a link from a known page to a new page,” and “when you submit a list of pages (a sitemap) for Google to crawl.”

Link-following is the primary mechanism; sitemaps are supplemental, not a substitute. Google’s own link best-practices guidance is blunt: “Every page you care about should have a link from at least one other page on your site.” An orphan with no sitemap entry and no external backlinks has none of the three discovery paths — it is genuinely invisible to Googlebot.

Evidence for this claim Google discovers pages through links and recommends that every page you care about have a link from at least one other page on the site. Scope: Current Google link and discovery guidance. Confidence: high · Verified: Google Search Central: Link best practices

John Mueller has called internal linking “super critical for SEO… one of the biggest things that you can do on a website to kind of guide Google and guide visitors to the pages that you think are important.” His framing is about link distance from your important hub pages, not just raw link count — an orphan has infinite link distance, because it isn’t reachable from your link graph at all.

They receive no internal PageRank

PageRank still underpins Google’s link-based authority, and internal links are the plumbing that moves it around your site. An orphan page is disconnected from that plumbing, so it receives zero internal PageRank no matter how good its content is.

This is the angle most people miss, and it cuts both ways. A page with strong external backlinks but no internal links isn’t just under-powered — it’s a PageRank leak. Authority flows into that page from the outside and then has nowhere to go, because no internal links carry it onward to the rest of your site. Add internal links and you both strengthen the orphan and let its inbound equity circulate.

They cluster in “Discovered — currently not indexed”

When Google knows a URL exists (from a sitemap or an external link) but decides not to index it, GSC reports “Discovered — currently not indexed.” Orphan pages are a classic cause: Google found the URL, but with no internal links pointing at it, there’s no signal that the page matters, so it gets deprioritized. Adding internal links from relevant pages is one of the most reliable ways to nudge those URLs out of that bucket. (It’s not the only cause — thin content and crawl budget play in too — but it’s one of the first things I check.)

They can waste crawl budget — but mostly on large sites

For the average site this barely registers. On very large sites (think 100k+ pages, or large numbers of rapidly changing URLs), big pools of orphaned and near-orphaned URLs consume crawl budget without earning their keep. Most sites genuinely don’t need to think about this; if you’re below that scale, treat orphans as a discovery and PageRank problem, not a crawl-budget one.

Common causes

  • Migrations and redesigns — URLs survive the move but the links that pointed to them don’t get rebuilt.
  • Navigation changes — a page gets pulled from the menu and nothing replaces the link.
  • CMS workflow gaps — a post is published without being assigned to a category or tag, so the only link path (the category archive) never gets created. WordPress and similar CMSes generate a lot of pagination and archive URLs that can strand content this way.
  • Discontinued products — the product page is left live after it’s removed from listings.
  • Campaign and PPC landing pages — built to be reached from an ad, never linked from the site (often intentional).
  • Staging, test, and siloed microsite content — left publicly accessible and never wired into the main link graph.

How to find orphan pages

The core method is a two-list comparison, because no crawl can find orphans on its own:

  1. List 1 — reachable pages. Run a site crawler that starts at your homepage and follows internal links. This is everything Google could discover by link-following.
  2. List 2 — all known pages. Gather every URL you know exists: your XML sitemap, Google Analytics (pages with sessions), GSC pages, your backlink index, and server logs.
  3. Orphans = in List 2 but not List 1.

In Ahrefs Site Audit, this is built in — it seeds the crawl from your sitemaps, the Ahrefs backlink index, and connected GA/GSC data, then flags any URL it knows about but couldn’t reach via internal links as an “Orphan page.” In Screaming Frog, you upload your sitemap, GA, and GSC data as URL sources, run the crawl in list/crawl mode, then use the crawl analysis to surface URLs that appear in those sources but weren’t found by link-following. With GSC alone you can’t see site structure, but URLs sitting in “Discovered — currently not indexed” that don’t appear in your crawl are strong orphan suspects.

How to fix them

Triage by value first — not every orphan deserves a link.

Before you triage, confirm the candidate itself is worth linking to: check its response status, its canonical target, and whether it’s set to noindex. Linking to a URL that redirects, 404s, canonicalizes elsewhere, or is deliberately excluded from the index isn’t a fix — it just adds a broken or wasted link. Only pages that resolve cleanly and are meant to be indexed belong in the “add links” bucket below.

  • Valuable pages you want to rank → add internal links. Find topically relevant pages and link from them with descriptive anchor text. Prioritize linking from pages that already have internal authority (hub and cluster pages), and make sure the page is in your XML sitemap too.
  • Intentional orphans (PPC, thank-you, some legal pages) → noindex. These don’t need internal links; they need to be kept out of the index cleanly. Adding links would be the wrong fix. Give each category of intentional orphan a named owner and a documented reason it’s excluded, so the next audit doesn’t re-flag it as an accidental one.
  • Low-value pages → redirect or delete. 301 to the most relevant alternative, or remove it entirely if there’s nothing relevant and no backlinks worth preserving.

Prevent recurrence: bake an internal-linking step into your publishing workflow, assign categories/tags before publishing, and re-audit after every migration, redesign, or navigation change. This is the same internal-link discipline that governs crawl depth and how PageRank flows through your architecture — orphan prevention is just the floor of good internal linking.

A note on the related-but-different term: an orphan page has no incoming internal links; a dead-end page has no outgoing internal links. They’re separate issues, and a page can be both.

Add an expert note

Pin an expert quote

New person? Create their unclaimed profile at /admin/experts/ → Pin an expert quote first.