Orphan and buried pages: finding the pages your own site forgets
An orphan page is a page that no other page on the site links to. A buried page is linked, but only from somewhere so deep that few readers or crawlers ever reach it. Both are common on sites that grew over years, and both waste content that was expensive to write.
On this page — 4 sections
Why do orphan and buried pages matter?
Quick answer
Google's link guidance says every page you care about should be linked from at least one other page on your site. Without internal links, a page is hard to discover, its relationship to the rest of the site is unclear, and readers never arrive from your other pages.
A Jeddah clinic might have a well-written page on children's fluoride treatment that was linked only from a 2022 newsletter post, since removed. The page still exists and may still be indexed, but nothing on the site leads to it.
How do you find orphan pages?
Quick answer
Compare two lists: every URL you know exists, from the sitemap and Search Console, and every URL a crawl reaches by following links from the home page. URLs in the first list but not the second are orphans, or are linked only from pages the crawl couldn't reach.
- Export known URLs from the sitemap and from Search Console's Performance report.
- Crawl the site from the home page, following internal links only.
- Subtract: known URLs the crawl never reached are your orphan list.
- Check each one: decide whether it should be linked, merged or removed, using the same keep, update, merge or remove decisions as the rest of the audit.
On an online store, orphans often turn out to be old campaign pages and products removed from categories but never from the site. On a content site, they are usually posts that fell off the paginated archive.
How do you find buried pages?
Quick answer
Use the crawl's click depth: how many clicks each page sits from the home page. Pages you care about that sit deeper than three or four clicks, a common practitioner threshold, are buried. Pages reachable only through paginated archives are the usual suspects.
An Egyptian recipe site may find that a popular kunafa recipe sits seven clicks deep: home, recipes, page 2, page 3 and so on through the archive. A link from a desserts hub would bring it to two clicks.
How do you reconnect them?
Quick answer
Give each page a parent in the hierarchy, link to it from that parent and from two or three related pages, and link from it back up. A hub that lists all its children fixes most buried pages at once, because it shortens every path.
- Assign a hub: the fluoride page belongs under the children's dentistry hub.
- Link from the hub with descriptive anchor text, such as "fluoride treatment for children".
- Link from related pages: the first-visit guide and the tooth-decay article both mention fluoride.
- Link back up: the fluoride page links to its hub near the top.
Then re-crawl. If the page now appears within a few clicks of the home page and has several internal links pointing to it, it has rejoined the site.
This article is part of the Content Operations series — Running content week to week: briefs that prevent rework, checks before publishing, fixing blockers first and refreshing on schedule.
About the author
Mohamed Youns
Semantic SEO Engineer · Author & system developer
Mohamed Youns writes about how search engines understand content — the same standards he applies when building semantic systems at Nut Hub. nut-hub.org