Skip to content

Orphan Page Checker

This orphan page checker fetches a sitemap, crawls up to 25 of the pages it lists, and reports which of those pages none of the others link to. An orphan page might still get traffic today, but with no internal links pointing at it, search engines have a harder time finding it, re-crawling it, or judging it important enough to rank.

Note

We fetch the sitemap server-side, then crawl up to 25 same-host listed URLs and extract each page's outgoing links. A URL is flagged as an orphan when none of the other crawled pages link to it: a self-link doesn't count, and this checks the sample against itself, not against your whole site.

Email me this result

Want a copy for your records or a client? Leave your email and we'll send you this result. It is written by a real person at Crawl Cove, not an automated blast, so it may take a little while to arrive. We'll only use your address to send it, unless you also tick the box below. Privacy policy.

That email address doesn't look right. Please check it and try again.

Something went wrong saving your address. Your result above is unaffected; please try again.

Got it. We'll email your result to you. Nothing else lands in your inbox unless you ticked the tips box.

What this orphan page checker shows you

Paste a sitemap URL and it is fetched and parsed server-side, the same way as our XML sitemap checker. Up to 25 of the listed URLs (on the same host as the sitemap) are then crawled: each page's HTML is fetched and its outgoing links extracted. A sitemap URL is flagged as an orphan when none of the other crawled pages link to it, a self-link does not count, and a URL blocked by robots.txt or one that redirects is reported but not crawled for links.

Why an orphan page is a problem even though it still exists

Search engines mostly discover and re-crawl pages by following links, not by re-reading a sitemap from scratch on every visit. A page with no internal links pointing at it is easy to miss on the first crawl and, once indexed, gets crawled less often than pages that sit in the middle of the link graph. That means slower re-indexing after you update it, and a weaker signal that the page matters at all: internal links are one of the ways search engines judge a page's importance relative to the rest of the site.

The one-sample limit, stated plainly

This checks whether the 25 (or fewer) pages crawled link to each other, not whether any page anywhere on your site links to them. A page outside the sample that links to one inside it is invisible to this check, so a real orphan on a larger site can be missed, and the result is a lower bound on how many orphans exist, not an exact count. A full crawl doesn't have that ceiling, because every page it finds is checked against every other page.

One sample here, your whole site there

Checking a capped sample is useful for a quick read, but the complete question, every indexable page against every other page, needs a real crawl. Crawl Cove builds the internal link graph as it crawls, flags every indexable page nothing else links to, and never flags the page you started the crawl from or lets a self-link rescue a page that nothing else points at.

Orphan page finder (whole-site) · Full site audit

Frequently asked questions

What is an orphan page?
A page that exists and may even be indexed, but has no internal links from other pages on the site pointing at it. Search engines mostly discover and re-crawl pages by following links, so an orphan page is easy to miss and slower to get re-indexed after changes.
Why does a self-link not rescue a page?
A page linking to itself, in its own navigation or a "you are here" breadcrumb, does not help search engines discover it from anywhere else on the site. This checker only counts a link from a different crawled page.
What does "blocked by robots.txt" mean in this result?
The page's path matches a Disallow rule in the site's robots.txt for all crawlers, so it was reported but not fetched or crawled for outgoing links, the same as a well-behaved search engine crawler would treat it.
Why does a redirected URL show no inlink count?
A URL that redirects was not fetched for its own content, so its outgoing links were not read. It is reported as redirected rather than checked as an orphan.
How many URLs does this check?
Up to 25 from the sitemap (or pooled across up to 5 child sitemaps, if it is a sitemap index), and only URLs on the same host as the sitemap itself. Each is checked against the others in that same sample, not against your whole site.
Is this orphan page checker free?
Yes. It is completely free and needs no sign-up.

From one page to your whole site

This checks one sample of up to 25 pages against each other. Crawl Cove builds the full internal link graph as it crawls your whole site, so every orphan is checked against every page rather than a capped sample, and rarer cases like sections cut off from the main link graph are still caught. See find orphan pages across your whole site or compare the plans.

Try it free

More free tools