Orphan Pages: How to Find Them and Fix Your Internal Linking

An orphan page is a URL on your site that no other page links to. It may sit in your XML sitemap, it may have been shared on social media once, but if you followed every link starting from the homepage you would never reach it.
Search engines can still find orphan pages through sitemaps and external links, so being orphaned does not mean being invisible. It does mean the page is missing one of the strongest signals a site can give: "this page is part of our content, and here is how it relates to everything else." For new sites and blogs where many pages struggle to get indexed, fixing orphans is one of the most practical improvements available.
Why orphan pages underperform
Internal links do three jobs at once.
Discovery. Crawlers follow links. A page that is only in the sitemap gets discovered, but it is not reinforced every time a related page is crawled.
Importance. The number and placement of internal links pointing to a page tell search engines how much the site values it. A page with zero links is effectively ranked last by its own site.
Context. Anchor text and the surrounding paragraph describe what the destination is about. Google's documentation on making links crawlable recommends descriptive anchor text for exactly this reason.
An orphan page gets none of these. It is common to see orphaned articles in Search Console under "Crawled – currently not indexed", especially on sites where most other pages are linked only through a chronological feed.
How orphan pages happen
Very few people create orphans on purpose. They tend to appear through a handful of predictable routes:
Feed-only navigation. Blogs that link to articles only from the homepage feed and paginated archives. Once an article drops off the first few pages, it is linked from nowhere that crawlers visit often.
No in-content links. Writers never link to related articles inside the body text. This is surprisingly common on multi-author platforms, where each contributor writes in isolation.
Site migrations and redesigns. Menus and footers change, and pages that were reachable through old navigation lose their only link.
Landing pages built for ads. Campaign pages intentionally left out of navigation, then later expected to rank.
Links that crawlers cannot follow. Navigation built with JavaScript click handlers rather than real
<a href>elements, or links markednofollowby a CMS default.
The last point deserves attention. Some editors and CMS templates add rel="nofollow" to every link inserted in the editor, internal ones included. That is a reasonable default for outbound links on user-generated content, but applied to internal links it tells search engines not to pass signals through your own navigation. If you run a site where contributors add links through a rich-text editor, inspect the published HTML of a few articles to confirm internal links are not being nofollowed.
Finding orphan pages
The method is the same regardless of tool: compare the list of URLs you want indexed with the list of URLs a crawler can reach by following links. Anything in the first list but missing from the second is an orphan.
Step 1: Build the "should exist" list
Combine your XML sitemap URLs, the pages listed in Google Search Console's Page indexing report, and, if available, pages that received organic landings in your analytics over the past year. Deduplicate and remove URLs that redirect or return errors.
Step 2: Build the "reachable" list
Run a crawler starting from the homepage, following only links. Desktop crawlers such as Screaming Frog SEO Spider and Sitebulb do this well, and most cloud SEO suites have an equivalent site audit. Make sure the crawl respects nofollow the way search engines do, and that it renders JavaScript if your navigation depends on it.
Step 3: Compare
Several crawlers can ingest the sitemap and analytics lists directly and report orphans automatically. If not, a spreadsheet VLOOKUP or MATCH across the two lists works fine for a site with a few hundred pages.
Step 4: Check the "weakly linked" pages too
A page linked only from page 9 of a category archive is technically not an orphan, but in practice it is close. Sort the reachable list by crawl depth (clicks from the homepage) and by number of internal links pointing in. Pages at depth 5 or deeper, or with only one or two inbound links, belong on the same to-do list.
Deciding what to do with each orphan
Not every orphan deserves rescuing. Sort them into three groups before you start adding links.
Situation | Action |
|---|---|
Useful page that fits the site's topics | Add contextual internal links from related pages |
Thin, outdated, or duplicate page | Merge into a stronger page and redirect, or remove it |
Intentionally isolated page (ad landing page, thank-you page) | Leave unlinked, add |
Linking to weak pages just to remove them from an orphan report spreads link signals onto content that will not rank anyway. Clean up first, then link.
Building internal links that actually help
Start with the pages that already perform
Look in Search Console for pages with the most impressions or clicks in the same topic area as the orphan. A link from a page Google already crawls frequently and values highly is worth more than several links from other weak pages.
Link where the reader would want to go next
The best internal link sits in a paragraph that naturally raises the question the destination answers. For example, an article about website speed that mentions image compression is a natural place to link to a detailed guide on image optimization. A "Related posts" widget at the bottom of the page is useful, but it is no substitute for in-context links.
Write descriptive anchors
"Our guide to cash flow forecasting" tells both readers and search engines what to expect. "Click here" and "this post" tell them nothing. Vary the wording naturally rather than repeating an exact-match keyword on every link.
Keep the number sensible
There is no magic count. For a typical 1,500-word article, three to five relevant internal links is a comfortable range. If you find yourself adding links that the reader would never click, stop.
Build hub pages for core topics
If you have ten articles about SEO fundamentals, a single overview page that introduces each and links to it gives every one of them a strong, permanent inbound link and gives readers a sensible starting point. Category pages can play this role if you add a short editorial introduction and order the articles deliberately rather than by date alone.
Make it part of the publishing routine
Fixing orphans once does not prevent new ones. Two habits keep the problem from coming back:
Link out when you publish. Every new article should link to two to five related existing articles where it genuinely helps.
Link in after you publish. Within a day or two of publishing, add a link to the new article from at least two older, related, already-indexed pages. This step is the one most sites skip, and it is the one that helps the new page most.
Run an orphan check every quarter, or after any redesign or migration. If you are also cleaning up after a traffic drop, the spam update recovery checklist includes internal linking alongside the other audits worth running, and the overview of why small business websites fail at SEO explains how structure fits with content and technical basics.
A note on JavaScript navigation
If your site renders menus, related-post widgets, or pagination with client-side JavaScript, confirm that the rendered HTML contains real <a href> links. Use the URL Inspection tool's "View tested page" option to look at what Google rendered. Links that only exist as click handlers, or that load after a user interaction, may not be followed. The article on JavaScript SEO and rendering delays goes deeper into this.






Comments (0)
No comments yet. Be the first to comment!