SIT-01
Orphan pages (no internal links)
What the check measures
A page gets this finding when all three conditions hold at once: it is listed in sitemap.xml, it returned a live response (2xx), and no other crawled page links to it. Such a page is called orphaned: it exists, but no route leads there. The address is compared in full, query parameters included: / and /login/ are two different addresses and only one of them can be in the sitemap. Until 19 September 2026 the parameters were dropped, and the finding could claim that an address was in the sitemap when it was nowhere in it.
The check is evaluated only on a full crawl, never on a sample. That is a deliberate limitation worth explaining: inbound links are counted only from pages we actually fetched. Over a sample of a few dozen pages the count would come out at zero for almost everything; the link usually comes from a page we never reached. Telling you “nothing links to this page” would then be a falsehood about your site caused by our limitation.
What the check does not do:
- It cannot see pages outside the sitemap. An orphaned page that is not in the sitemap either gets no finding; we have no way to find it. It is a blind spot we have nothing to fill with.
- It does not count inbound links from other sites. A page linked by half the internet but not by your own site is orphaned as far as this check goes.
- It does not care where the link comes from. A single footer link counts the same as one from the main navigation.
- It does not judge whether a page should be orphaned. Campaign landing pages, order confirmations and test URLs are deliberately unlinked.
- It cannot see links rendered by JavaScript unless the page was sampled through a browser (
ACC-03).
The finding is listed per page so you know which ones they are. Severity is warning.
How strong the evidence is
We recommend it because it does no harm or has some other benefit, but we promise nothing about whether it makes language models cite you. Nobody has demonstrated that yet.
We have no evidence that linking a page up improves your chances of being cited in an AI answer. Hence the “effect not demonstrated” class.
And there is something else we have to admit, because it weakens the finding's own logic: the page this finding talks about is, by definition, in the sitemap. A sitemap is, per Google's documentation, precisely the mechanism for telling a search engine about URLs it might not otherwise reach. So this is not a page nobody can find; it is a page findable only that way.
So why report it? For two reasons, both about your site rather than about an algorithm:
- An orphaned page is usually a symptom, not the disease. When no route leads to content from your navigation, it is nearly always because it was forgotten: an old campaign, content stranded by a site rebuild, a page published outside the normal process. The finding reminds you.
- A sitemap does not guarantee indexing on its own. Google says so literally: it does not guarantee that all the items in your sitemap will be crawled and indexed. A link from your own site is a second, independent route to the page, and two beat one.
What we do not claim: that internal linking boosts a page's “authority” or moves it up the rankings. Domain authority, backlinks and domain age have no documented support in the context of AI visibility, and we deliberately keep them out of our score. It is a classic SEO claim transplanted into GEO without evidence.
In practice: go through the list and ask of each page whether it belongs there. If it does, add a link. If it does not, remove it from the sitemap; that clears the finding correctly, rather than by a workaround.
How to fix it
A list of orphaned pages almost always breaks into three groups, each handled differently:
| What it is | What to do |
|---|---|
| Content meant to be reachable: an article, a product, a service | Add a link from wherever a visitor would look for it: a hub page, a category, a related article. |
| A page meant to stay out of navigation: a campaign landing page, an order confirmation | Remove it from the sitemap. If it does not belong in search at all, give it noindex too (ACC-07). |
| A leftover from a site rebuild: an old URL, a duplicate version | Redirect it permanently (301) to the current content and drop it from the sitemap. |
When adding links, two things matter more than their number:
- The link must be a real
<a href>. An element that calls a script on click and changes the address is not a link; no crawler will follow it and it will not appear in the link graph. - The link text should describe the destination. “Heat pump servicing” is a link; “click here” is a wasted chance to say what is on the other side.
And one systemic note that saves repeat work: orphaned pages appear where content is published outside the template. When every article is automatically linked from its category listing and every product from a product list, orphans barely occur. If your list is long, look for the cause in your publishing process, not in individual pages.
What the report says about it
Finding description
The page is listed in sitemap.xml, but none of the pages this audit visited links to it internally. If the audit did not cover the whole site, the link may sit on a page we never visited. Without an internal link, the page is hard to discover for both regular visitors and crawlers that follow links rather than only the sitemap.
Recommendation
Add at least one internal link to this page somewhere relevant on the site (navigation, related articles, hub pages). If the page is no longer current, consider removing it from sitemap.xml and setting up a redirect.
Sources
- Google Search Central: Learn about sitemaps (accessed 2026-09-12)
- Google Search Central: SEO Starter Guide (accessed 2026-09-12)
Text verified 2026-09-19