IndexVexa
Free SEO Tools

Orphan Page Checker

Enter a website to find potential orphan pages — URLs listed in the sitemap but not discovered through internal links during a bounded crawl.

We always start from the site's homepage and crawl up to 50 internal pages — this can take up to half a minute.

What is an orphan page?

An orphan page is a URL that exists in a site's sitemap but isn't reachable through the site's own internal links. Search engines and visitors alike rely on internal links to find pages — a page with no internal links pointing to it can be harder to discover, even if it's listed in the sitemap.

This tool fetches robots.txt, discovers the sitemap, crawls the site from its homepage following internal links up to a bounded limit, and compares what the sitemap lists against what the crawl actually found. Any crawl is necessarily limited — this tool cannot prove a page is definitively orphaned, only that it wasn't found within this check's bounds. Results are labeled “Potential Orphan Page” for that reason.

Frequently Asked Questions

Why does it say "potential" orphan, not just "orphan"?
The crawl is bounded — up to 50 pages — so it can't follow every link on a large site. A page not found by this crawl may still have internal links pointing to it elsewhere on the site.
Why do some sitemap URLs show as "Not analyzed" instead of orphan?
When the crawl stops at its page or time limit before covering the whole site, any sitemap URL it hasn't reached yet is genuinely unknown, not orphaned — it's labeled "Not analyzed" rather than guessed at.
Does this tool ignore robots.txt?
No. Pages disallowed by robots.txt for general crawlers are not fetched during the crawl, which can make results incomplete on sites with restrictive robots rules.