An orphan page is a page on your website that no other page on the same site links to. It exists and may even be indexed, but there is no path to it through your menus, content or footer, so visitors cannot find it by browsing and search engines only reach it through a sitemap or a link from somewhere else.
How orphan pages happen
Search engines discover most pages by following links. A crawler lands on your homepage, follows every link it finds, then the links on those pages, and so on. This is crawling, and it is how Google maps the shape of a site. A page with no internal links pointing to it sits outside that map.
Orphans are rarely created on purpose. The usual causes are ordinary site housekeeping:
- A menu is redesigned and an old section drops out of it, while its pages stay live.
- A landing page is built for an ad campaign and never linked from the main site.
- Blog posts slide off the paginated archive and are linked from nowhere else.
- Products are removed from categories but their pages remain published.
- A migration changes URLs, the internal links are not updated, and the old addresses now return errors or redirect somewhere else.
Pages listed in your XML sitemap can still be found and indexed, so an orphan is not always invisible. But a page the site itself never links to gives Google little reason to think it matters, and it receives none of the authority your other pages could pass on.
Why it matters
Internal links do three jobs at once: they help people move around, they help crawlers find pages, and they pass authority from strong pages to weaker ones. An orphan gets none of that. If the page is useful, such as a service page for a borough you cover or a guide that answers a common customer question, it is underperforming for a reason that is cheap to fix.
If the page is not useful, the problem is the opposite. Forgotten orphans often include old offers, test pages, duplicate drafts and expired event pages. They can still appear in search, show outdated prices or information, and add to the pile of low-value URLs that Google has to crawl.
Common mistakes
- Relying on the sitemap to carry important pages. A sitemap tells Google a page exists; it does not tell Google the page matters.
- Linking every orphan from the footer. A footer stuffed with links is a quick patch, not a structure. Links work best from related content.
- Treating every orphan as a page to save. Some should be redirected or removed instead.
- Forgetting ad landing pages. Paid campaign pages are often orphans on purpose. Decide whether they should be indexed at all and set a noindex if not.
How to act on it
First, find them. Run a crawl of the site with a tool such as Screaming Frog, which only finds pages it can reach by links. Then compare that list with every URL you know about from other sources: your XML sitemap, the pages receiving visits in GA4 and the pages showing impressions in Google Search Console. Anything that appears in those lists but not in the crawl is an orphan. Several crawlers will run this comparison for you once you connect the data sources.
Next, sort each orphan into one of three groups. Keep and link: useful, current pages that belong in your site architecture, which should get links from their parent category and from two or three closely related pages. Merge and redirect: pages that overlap with a stronger page, which should be combined and sent there with a 301. Remove: pages with no value, no traffic and no backlinks, which can return a 404 or 410.
Finally, prevent new ones. Add a step to your publishing routine: no page goes live without at least one link from a relevant existing page. Re-run the comparison every quarter or after any redesign. A crawl comparison of this kind is one of the first checks in my technical SEO work, because it often surfaces pages nobody on the team knew were still live.
