Orphan Page
Definition
A page becomes an orphan when it is published without a link from the navigation, a listing, a hub page or any other document on the same site. It is not the same as a page marked noindex, which you excluded deliberately, nor a page blocked in robots.txt, which the crawler is told to leave alone. An orphan is meant to be found and simply cannot be. Search engines discover pages mostly by following links, so an orphan surfaces only if it appears in a sitemap.xml, is linked from another site, or is submitted by hand. Even once indexed it receives no internal link signal at all, and pages in that state rarely rank for anything competitive.
Why It Matters
Orphans are quiet waste. The page was researched, written and published, and it earns nothing because the one cheap thing it needed — a link from somewhere — was never added. They gather in predictable places: retired campaign pages, PDFs uploaded for a single email, pages created by a tool that does not touch the navigation. The visible symptom is a URL sitting in Search Console under Discovered — currently not indexed for months on end. On a larger site they also eat crawl budget, because the crawler keeps re-checking addresses that nothing on the site treats as worth linking to.
How It Works
Finding them means comparing two lists. First, crawl the site with a tool that follows links only, which gives you the set of pages reachable from the homepage. Second, build the set of pages that exist, from the sitemap, the file listing on the host and the Pages report in Search Console. Anything in the second list and missing from the first is an orphan. The fix is nearly always a link: add the page to the relevant hub, to the navigation, or to a related-reading block on a page covering the same ground. Adding it to the sitemap alone does not count — that may get it crawled, but it gives the page no standing inside the site.
Real-World Example
A studio uploads a case study to 99helpers at bridge-rebuild.99helpers.site and sends the link to one prospective client. Nothing on the main site mentions it. A year on, the case study has had no search traffic whatsoever, because no page anywhere points at it. Adding it to the work index and linking it from two related write-ups gives it a route in, and it is indexed within a fortnight.
Common Mistakes
- ✕Counting a sitemap entry as a link — a sitemap lists addresses and passes no signal about importance or context
- ✕Linking only from a menu built by JavaScript after load — if the link is not in the rendered markup, the crawler may never follow it
- ✕Leaving one-off PDFs and campaign pages unlinked once the campaign ends — they are the most common orphans on any site
Related Terms
Internal Linking
Internal linking is the practice of linking from one page of a site to another page on the same site. It is how both readers and crawlers find their way past the home page.
Crawl Budget
Crawl budget is the number of URLs a search engine is willing and able to fetch from one site in a given period. For most sites it is not a real constraint.
Indexability
Indexability is whether a page is technically allowed into a search engine's index. It is eligibility, not a promise — an indexable page can still be left out.
sitemap.xml
An XML file listing the pages of a site so search engines do not have to find them all by following links. It aids discovery; it does not affect ranking.
Landing Page
A page built for one purpose, which people reach directly from an ad, an email or a shared link. It is judged by the share of visitors who do the one thing it asks.
Put a file online in seconds
Drop in a document, an image, a page or a whole static website and share the link — free, with no build step and no server to set up.
Host a file free →