> "Every internal link is a small sentence your site says to Google about which of its pages matter most.
Prerequisites
- 14
Learning Objectives
- Distinguish flat from deep site architecture and apply the crawl-depth ('clicks from home') heuristic without treating 'three clicks' as a law.
- Design descriptive, hierarchical, stable URLs — and judge honestly how much a URL matters for ranking versus for users.
- Explain how main navigation, breadcrumbs, and footers serve crawling and users at once, and where each helps or hurts.
- Describe how link equity — the internal analog of PageRank — flows through a site's internal links, and where it pools or leaks away.
- Point earned authority at the pages that need it with strategic internal linking, and add the internal links that finish the page stuck at #11.
- Choose between hub-and-spoke and silo structures, and see through the manipulative 'PageRank sculpting' and rigid-silo myths.
- Find and fix orphan pages and crawl traps, and explain why a page nothing links to is a page that was never in the race.
In This Chapter
- Overview
- Learning Paths
- 15.1 Flat vs. deep: the reach of every page
- 15.2 URL structure: descriptive, hierarchical, stable
- 15.3 Navigation: menus, breadcrumbs, and footers
- 15.4 Internal linking: how authority flows through a site
- 15.5 Strategic internal linking: pointing authority where it's needed
- 15.6 Hub-and-spoke and silos: organizing by topic
- 15.7 Orphan pages and crawl traps: the two failures that waste the most
- 📈 The Strategy File
- Conclusion
- Key Terms
- Spaced Review
Chapter 15: Site Architecture — How Your Pages Are Organized Determines How Authority Flows
"Every internal link is a small sentence your site says to Google about which of its pages matter most. Most sites are mumbling." — constructed for this book, in the narrator's voice
Overview
Here is a puzzle that will sharpen the rest of this chapter. Two pages on the same website are, on their own merits, about equally good — similar quality, similar length, similar topic. One ranks on page one and pulls steady traffic. The other is marooned, invisible, pulling nothing. Nothing about either page, read in isolation, explains the difference. So what does?
Very often, the answer is not on the page at all. It is where the page sits in the site — how many clicks it takes to reach, what links to it, and how much of the site's hard-won authority actually flows down to it. This is site architecture: the way your pages are organized and connected, from the shape of your URLs to the structure of your menus to the web of internal links that ties everything together. It is, page for page, the most underused lever in all of SEO — because it is invisible on any single page, it never shows up in the "optimize this article" checklists, and it pays off in a way that is real but diffuse. Most site owners never touch it. The ones who do find rankings they could not buy with content alone.
The reason architecture matters comes straight from the pipeline of Chapter 1. Google discovers pages by following links, so a page nothing links to may never be found. Google spends its crawling on the pages your structure signals are important, so a page buried six clicks deep gets crawled rarely and treated as an afterthought. And Google's oldest ranking idea — that links are votes of importance — operates inside your site as much as between sites: your internal links quietly tell Google which of your own pages you consider most important, and pass a share of your site's authority to them. Get the architecture right and you are helping Google at three stages of the pipeline at once. Get it wrong and you can strand your best work where no one, human or crawler, will ever find it.
This is the technical foundation the book's fourth theme is about — the ground content and authority stand on — and it connects directly to the fifth: links are the strongest signal off-page, and architecture is how you steer the authority those links create once it arrives.
In this chapter, you will learn to:
- Tell flat from deep architecture, and use "clicks from the homepage" as a diagnostic — while refusing the folklore that a magic click-number is a ranking law.
- Structure URLs that are descriptive, hierarchical, and stable, and weigh honestly how little a URL does for ranking and how much it does for users.
- Use main navigation, breadcrumbs, and footers deliberately — each of which serves the crawler and the human at the same time.
- Trace how link equity flows through internal links, where it pools, and where it leaks.
- Point authority from your strongest pages at the pages that need it — and finally close the last open gap on the page stuck at #11.
- Design a hub-and-spoke or silo structure without falling for "PageRank sculpting" or rigid siloing.
- Hunt down orphan pages and crawl traps, the two architecture failures that quietly waste the most.
Learning Paths
Architecture touches every kind of site, so weight this chapter to yours. 🏪 Local Business: §15.1, §15.5, and the Strategy File are your core — we design Rivertown's service-and-city structure and fix the orphaned service pages that keep vanishing. 📝 Content Creator: live in §15.4 and §15.6 — internal linking and the hub-and-spoke model are how a blog's authority stops leaking and starts compounding. 🛒 E-Commerce: §15.1 (depth), §15.2 (URLs), and §15.7 (crawl traps) are where category depth and faceted navigation make or break you; note the forward pointer to Chapter 31. 🔧 Developer: §15.2 (URL rules and redirects), §15.3 (nav and breadcrumb markup), and §15.7 (crawl traps and infinite spaces) are your sections — you are the one who implements them. 📊 Strategist: the whole chapter is a lever you will reach for in every audit; §15.4 and §15.5 turn "we have authority" into "our important pages have authority."
15.1 Flat vs. deep: the reach of every page
Start with a single, physical question about any page on your site: how many clicks does it take to get there from the homepage? That number is the page's crawl depth (also called click depth) — the count of link "hops" from your main entry point to the page in question. The homepage is depth 0. A page linked directly from the homepage is depth 1. A page you can only reach by clicking home → category → subcategory → product is depth 3. Crawl depth is the simplest, most useful measurement in all of site architecture, and almost nobody measures it.
From crawl depth comes the distinction that organizes this whole section. A flat architecture is one where every important page is reachable in a small number of clicks — the site is "wide," with most content near the surface. A deep architecture is one where reaching content requires many clicks down through layers of nested categories — the site is "tall," with content buried far from home. Here is the contrast, drawn as the crawler experiences it.
FLAT vs. DEEP ARCHITECTURE [schematic — not to scale]
FLAT (wide — good) DEEP (tall — risky)
HOME (depth 0) HOME (depth 0)
├─ Service page (depth 1) └─ Category (depth 1)
├─ Service page (depth 1) └─ Subcat (depth 2)
├─ Service page (depth 1) └─ Subsub (depth 3)
├─ Location page (depth 1) └─ Page (depth 4)
└─ Guide (depth 1) └─ deeper… (depth 5+)
Most pages sit 1–3 clicks from home. Important pages sit 4, 5, 6+ clicks down.
Googlebot reaches them easily; link Googlebot reaches them late and rarely;
equity flows down in a few hops. authority thins to almost nothing on the way.
Why does depth matter to a search engine? Two mechanisms, both from earlier chapters. First, discovery and crawl priority (Chapter 1, Chapter 14): Google generally crawls shallow pages more often and treats distance-from-home as a rough signal of how important you think a page is. A page you placed one click from your homepage is a page you are shouting about; a page you buried five levels down is one you are whispering about, and Google reads the whisper. Second, link equity (§15.4): authority spreads through links and thins as it travels, so a page far from your well-linked homepage receives a fainter share of it. Depth is not a penalty applied by a rule; it is the natural consequence of how discovery and link flow work.
🔎 How Search Sees It Google's Search Advocates have said, in public guidance and Q&A over the years, that how many clicks it takes to reach a page from the homepage matters more than how deep the page's URL path looks. That is a distinction worth holding: a page at the URL
example.com/a/b/c/d/pageis not "deep" in the sense that hurts you if it is linked directly from the homepage — the slashes in the URL are not the problem. What Google is estimating is link distance: how far the page is along the actual link graph from your most important, most-linked pages. A shallow link path with a long URL path is fine. A short URL that is nonetheless reachable only after six clicks is the real "deep" page. Measure clicks, not slashes.
This is where the famous rule enters — and where we defuse it.
🚫 SEO Myth: "Every page must be within three clicks of the homepage or it won't rank." The "three-click rule" began life as a usability heuristic — a rough idea that users get frustrated hunting more than a few clicks for what they want — and got imported into SEO as a hard law. It is not one. There is no magic number Google enforces; a page four clicks deep is not doomed, and a page one click deep is not thereby guaranteed anything. What is true underneath the myth is directional and worth keeping: your important pages should be easy to reach, and burying them deep makes them harder for both crawlers and humans to find. So use "clicks from home" as a diagnostic, not a commandment. If your money pages sit five and six clicks down, that is a real problem to fix. If a genuinely minor page sits four clicks deep, relax — depth is a symptom to investigate, never a threshold to obsess over.
There is a matching error in the other direction, and large sites fall into it: making everything as flat as possible, so the homepage or a single mega-menu links to hundreds or thousands of pages. That is not free. As §15.4 will show, a page that links to a thousand others passes each of them a vanishingly small share of its authority, and a menu with hundreds of links serves no human at all. Flatness is a virtue up to the point where it becomes a dumping ground. The real target is not "flat" or "deep" as a slogan; it is a logical hierarchy — related things grouped together, important things near the surface, nothing important stranded in the basement, and nothing so wide that structure dissolves into noise.
🛠️ Try It on Your Site You can measure your own crawl depth for free. The blunt version: click from your homepage to your single most important commercial page — the one that makes you money — counting clicks, using only links a normal visitor would use (not the address bar). If it took more than three or four, ask why. The thorough version: run a free crawler (the free tier of a desktop crawler will handle up to a few hundred URLs) and look at its "crawl depth" column; sort descending and read the deepest pages. You will almost always find something important that got buried by accident when the site grew. That page is your first architecture win.
15.2 URL structure: descriptive, hierarchical, stable
A URL (Uniform Resource Locator) is a page's address — the string in the browser bar. URL structure is the deliberate design of those addresses across a site: the words in them, the folder hierarchy they express, and the rules that keep them consistent. Good URL structure is one of those quiet disciplines that costs almost nothing when you build a site and is painfully expensive to retrofit later, which is exactly why it is worth getting right the first time.
Three qualities define a good URL. It is descriptive (a human can read it and know what the page is), it is hierarchical (its folders express where the page sits in the site's structure), and it is stable (it does not change, because a changed URL breaks every link and bookmark pointing at the old one). Compare:
URL STRUCTURE — GOOD vs. BAD [schematic]
BAD GOOD
─── ────
/index.php?id=48213&cat=7 /services/plumbing/water-heater-replacement/
→ opaque; tells human and Google nothing → readable; a person knows the topic before clicking
/Services/Plumbing/WaterHeater_REPAIR /services/plumbing/water-heater-repair/
→ mixed case + underscores; fragile → lowercase, hyphen-separated words, clean
/p/9f3a-x /guides/how-to-relight-a-pilot-light/
→ a random slug; unshareable → the slug IS the topic; shareable and self-describing
/water-heater-replacement-cost-guide-2024/ /guides/water-heater-replacement-cost/
→ a date baked in that will rot next year → evergreen; no year to go stale
The conventions that produce the good column are settled and uncontroversial: use lowercase letters; separate
words with hyphens, not underscores or spaces (Google has long recommended hyphens as word separators);
include the actual words that describe the page rather than opaque IDs; keep the path shallow and logical; and
avoid baking in anything that will change — a year, a campaign name, a technology. Hierarchy in the path is a
gift to both the reader and the machine: /services/plumbing/water-heater-replacement/ tells a human, at a
glance, that this is a plumbing service, and it mirrors the site structure we will design in the Strategy File.
Now for the honest part, because URLs attract more mythology per character than almost anything in SEO.
⚖️ Evidence Check Claim: "Keywords in the URL are an important ranking factor — pack your target keyword into the URL." — Confirmed by Google: words in the URL are a ranking signal, but a very small one. Google's Search Advocates have repeatedly described keywords in the URL as a minor, lightweight factor — something that plays "a tiny role," not a lever worth contorting your site for. That the signal exists is confirmed; that it is small is also, unusually, confirmed. — Strong practitioner consensus: the real value of a good URL is for humans — a readable URL earns marginally more clicks in the SERP, is more trustworthy, is easier to share, and is easier to build links to because people can see what they are linking to. Those human benefits dwarf the tiny ranking hint. — The trap: because "keywords in URL" sounds like a ranking factor, people over-optimize — stuffing
/best-cheap-affordable-water-heater-repair-service-near-me/and, worse, changing existing URLs to add keywords. Changing a URL to chase a tiny ranking hint, and thereby breaking every link and needing redirects, is one of the clearest bad trades in SEO. Get URLs right on new pages; almost never rewrite a working URL just to add a keyword. The juice is not worth the squeeze.
That last point deserves its own emphasis, because it is where URL enthusiasm does real damage. A URL that is already indexed and earning traffic is stable value. If you change it, every external link and internal link to the old address now points at a dead page unless you set up a 301 redirect (the permanent-move instruction we cover in Chapters 14 and 21) for every single one. Redirects work, but they are a maintenance burden and a small loss, and a redirect chain (a redirect pointing to another redirect) is a genuine problem. So the rule of thumb is asymmetric: lavish care on the URLs of new pages, where good structure is free, and treat the URLs of existing, working pages as close to sacred — change them only for a real reason (a genuine restructure, an HTTPS migration, a truly broken pattern), and when you do, map every redirect carefully.
🔗 Connection The mechanics of changing URLs safely — 301 redirects, redirect maps, and the discipline that keeps a migration from becoming a catastrophe — are the whole subject of Chapter 21 (Site Migrations), and the
robots.txtand canonical-tag tools that manage duplicate and parameter URLs live in Chapter 14. When URL parameters multiply out of control — the classic e-commerce faceted-navigation problem — that is Chapter 31. Here we care only about designing clean, stable URLs in the first place.
15.3 Navigation: menus, breadcrumbs, and footers
Navigation is the visible surface of your architecture — the menus and links a visitor actually uses to move around. It is also, and this is the point most people miss, one of the primary ways Googlebot moves around, because the crawler follows the same links your visitors click. Every navigational choice is therefore a choice about both audiences at once. Three navigational structures do most of the work.
Main (primary) navigation is the menu at the top of every page — the site's table of contents. Because it appears site-wide, every link in it is a link from every page, which makes main-nav links unusually powerful for both discovery and link flow: a page in the main menu is, in effect, one click from everywhere. This is exactly why the main navigation should point at your most important pages (your core service or category pages) and not try to list everything. A bloated mega-menu with two hundred links dilutes the signal and overwhelms the human. Curate it.
Breadcrumbs are the small trail of links, usually near the top of a page, that shows where the page sits
in the hierarchy — for example: Home › Services › Plumbing › Water Heater Replacement. A breadcrumb is a
secondary navigation aid that names the page's ancestors and lets a user jump back up a level. Breadcrumbs are
a genuinely elegant piece of architecture because they do four useful things at once, drawn here:
WHAT A BREADCRUMB DOES [schematic]
Home › Services › Plumbing › Water Heater Replacement
│ │ │ │
│ │ │ └─ (you are here — not a link)
│ │ └─ up one level: all plumbing services
│ └─ up two levels: all services
└─ back to the top
FOR THE USER: orientation ("where am I, and how do I get back up?")
FOR THE CRAWLER: extra internal links UP the hierarchy, reinforcing structure
FOR LINK FLOW: passes equity back to parent/category pages on every child page
FOR THE SERP: can render as a breadcrumb trail under your title (with schema — Ch 18)
Notice the third line: because a breadcrumb appears on every child page and links up to the category and home, it quietly sends a share of authority from your many deep pages back up to your important hub pages — reinforcing them. Breadcrumbs are one of the few structures that improve orientation, crawlability, and link flow simultaneously, which is why they are close to a free win on any hierarchical site.
🔗 Connection Breadcrumbs can also earn a visible breadcrumb trail in your search result — the little
example.com › Services › Plumbingline under some listings — but only when you mark them up with breadcrumb structured data. That markup, and the rest of the schema that makes your structure explicit to Google, is Chapter 18 (Structured Data and Schema Markup). The breadcrumb as architecture belongs here; the breadcrumb as a rich result belongs there.
Footer navigation is the block of links at the bottom of every page. Footers are useful for genuinely sitewide, secondary links — contact, about, privacy, main categories, key locations — but they are also the single most abused piece of navigation in SEO. The abuse pattern: cramming the footer with dozens or hundreds of keyword-stuffed links ("water heater repair Cedar Hills," "water heater repair Northgate," "furnace repair Cedar Hills," and on and on) in the belief that these sitewide links will push those pages up. They mostly do not, they make the footer a wall of spam that helps no human, and at their worst they look exactly like the manipulative sitewide-link patterns Google has learned to discount. A footer is for a handful of important, genuinely useful sitewide links — not a dumping ground for every page you wish ranked.
🔄 Check Your Understanding Your homepage's main navigation currently lists 40 items, including every individual service in every city. Your most important page — the main "Plumbing Services" hub — is buried as the 27th item in a dropdown. Name two distinct problems this creates, one for humans and one for link flow.
Answer
For humans: a 40-item menu is unscannable — no visitor can quickly find the plumbing hub, so navigation fails at its one job (orientation and movement). For link flow: because the menu appears sitewide, it makes all 40 pages one click from everywhere, spreading the homepage's authority thinly across 40 destinations instead of concentrating it on the few hubs that matter — and it buries the important plumbing hub among trivial links, giving Google no signal that it is more important than the other 39. The fix: curate the main nav to a handful of top-level hubs, and let those hubs link down to the specifics.
15.4 Internal linking: how authority flows through a site
We now reach the heart of the chapter, and the most underused lever in SEO. An internal link is a link from one page on your site to another page on the same site. That is the whole definition, and it sounds almost too mundane to matter. It is, in fact, one of the most powerful tools you have — because internal links do two large jobs at once: they help Google discover and understand your pages, and they distribute authority among them.
The discovery job we already met in Chapter 1: Google finds pages by following links, so internal links are
the roads that let the crawler reach your content, and the anchor text of an internal link (the visible,
clickable words) tells Google what the destination page is about. A link that says "read our
water heater replacement cost guide" tells Google far more than one that says "click here." That is the
understanding half.
The authority half is where internal linking becomes strategic, and to explain it we need the idea that built Google. In 1998, its founders described a model called PageRank: treat every link on the web as a vote of importance, so that a page accumulates authority from the pages that link to it, and — crucially — a page passes a share of its own authority along through its outgoing links. A vote from an important page counts for more than a vote from an obscure one. We give PageRank its full treatment, and its one honest formula, in Chapter 22; for now we need only its consequence for your own site. Because your internal links are also votes, the authority your site earns from the outside world does not sit still — it flows through your internal link graph, from page to page, pooling where many links point and thinning where few do.
The name for that flowing authority, as it moves between your pages, is link equity: the ranking value that a link passes from the source page to the target page. (You will also hear the informal "link juice" — same idea, less dignified.) Link equity is the internal expression of PageRank. It is what you are steering when you decide which pages link to which. And here is the mechanism that makes internal linking a genuine lever rather than a decoration:
HOW LINK EQUITY FLOWS INTERNALLY [simplified illustration — real model is Ch 22]
The HOMEPAGE has the most authority (most external links point at the front door).
It passes a share of that authority through each of its outgoing links.
HOMEPAGE (authority ≈ 100, illustrative)
│
│ links to 5 pages → each receives roughly 100 ÷ 5 = ~20
├──────────▶ Plumbing hub (~20) ──┐
├──────────▶ HVAC hub (~20) │ each hub then passes
├──────────▶ Electrical hub (~20) │ ITS share down to its
├──────────▶ About (~20) │ own child pages…
└──────────▶ Contact (~20) ──┘
A page that links to 5 pages passes each ~1/5 of its equity.
A page that links to 100 pages passes each ~1/100 — the same authority, split 20× thinner.
So WHERE your strong pages link, and HOW MANY links they spread it across, decides
which of your pages get the authority to compete.
Two rules fall out of that picture, and both are load-bearing. First, a link from a page that links to few things passes more equity than a link from a page that links to hundreds — the authority is divided among the outgoing links, so fewer links means a larger slice each. This is the real reason a two-hundred-link footer is weak: each of those links carries a two-hundredth of the page's equity. Second, equity thins as it travels: a page three hops from your homepage receives far less than one linked directly from it, which is the §15.1 depth point restated in link terms. Put the two together and the strategy writes itself — concentrate your important pages near your strongest pages, and don't drown your strong pages' links in noise.
⚖️ Evidence Check Claim: "Internal linking actually moves rankings." Where does this sit on our honesty scale? This one is unusually well-supported, so it is worth sorting carefully. — Confirmed by Google: Google's own documentation states plainly that internal links help Google find pages and understand a site's structure and the relative importance of its pages, and Google's Search Advocates have repeatedly affirmed that internal linking is genuinely important and often underused. That internal links matter for discovery, understanding, and importance is confirmed — a rarity in a field where Google confirms little. — Strong correlation and professional experience: that strategic internal linking (pointing authority at a specific target page) moves that page up is supported by consistent practitioner results and vendor studies — I have watched a page climb after nothing changed but the internal links pointing to it. But note the honesty: this is professional experience and correlation, not a published magnitude. — What no one can give you: a number. Google publishes no weight for internal links, and anyone who tells you "internal links are X% of ranking" is inventing it. We know the direction is real and confirmed; we do not know the size, and we say so.
The honest limit, and it matters for the theme running under this chapter: internal links redistribute the authority you have earned — they do not manufacture it from nothing. You cannot conjure a strong site out of a weak one by linking your own pages together in clever patterns; there is no self-referential loophole where a site with no external authority link-builds itself to the top from the inside. What internal linking does is steer the authority your genuine external links (Chapters 22–24) bring in, directing it to the pages that most need to compete. That is the boundary between a legitimate lever and a trick, and it is precisely theme five: links are the strongest off-page signal, and internal linking is how you make the most of the ones you have honestly earned — not a substitute for earning them.
15.5 Strategic internal linking: pointing authority where it's needed
Once you understand that link equity flows, internal linking stops being "add some related links" and becomes a deliberate act: find your strongest pages, and point them at the pages you need to lift. This is the single highest-leverage internal-linking move, and it is almost entirely neglected because it is invisible on the target page — you improve a page by editing other pages.
The method is straightforward:
- Find your strongest pages. Your homepage is usually one. Beyond it, the pages with the most external links and the most traffic hold the most authority — often an older cornerstone guide, a popular blog post, or a well-linked resource. (You will identify these precisely in Search Console's Links report and your analytics; see the box below.)
- Find your targets — pages that deserve to rank but are underpowered: a new page with no links yet, or a genuinely good page stuck just off page one.
- Add relevant internal links from the strong pages to the targets, with descriptive anchor text, where it genuinely helps the reader. Relevance is the constraint that keeps this honest: you link because a reader of the strong page would genuinely want the target, not merely because you want to shove equity around.
Which brings us, at last, to a page that has been waiting since Chapter 1 for exactly this fix.
You met Rivertown's water-heater guide — "the page stuck at #11" — in Chapter 1, where Figure 1.3 diagnosed why it was stranded at the top of page two: a title that answered "who we are" instead of the searcher's question, a page thinner than the result outranking it, and — the gap this chapter owns — not one other page on the site linked to it. In Chapter 3 you diagnosed the intent mismatch; in Chapter 9 you fixed the title, the meta description, the heading structure, and the comprehensiveness, and then deliberately deferred the internal-links fix to here. This is here. Let's close it.
📄 Read the Report — finishing the #11 page
text FIGURE 15.1 — "The internal links that finish the water-heater page" [the Strategy File] THE QUERY / PAGE Rivertown's water-heater replacement guide (the #11 page), AFTER its Ch 9 on-page rewrite — but still receiving ZERO internal links from the rest of the site. WHAT'S THERE Before: the page is an island. The homepage doesn't link to it; no service or blog page links to it; it sits in the sitemap and nowhere else. It receives essentially no internal link equity, and Google gets no internal signal that Rivertown considers it important. After — we add relevant internal links FROM pages that already have authority: • the new "Plumbing Services" hub → links down to it (its natural parent) • the homepage's "Popular Guides" module → links to it (homepage authority) • 3 related blog posts already getting traffic ("signs your water heater is failing," "tank vs. tankless," "how to flush a water heater") → each links to it with descriptive anchor text, where a reader would genuinely want it • the breadcrumb on the page links UP to the plumbing hub, reinforcing the cluster WHAT IT SHOWS The page now receives equity from the homepage and from several trusted pages, and the anchor text ("water heater replacement cost," "replacing a water heater") tells Google what it is about. Internally, Rivertown is now voting for this page. WHAT IT DOESN'T Internal links alone do not guarantee page one, and they do not add EXTERNAL authority (that is Chapters 22–24). We are redistributing earned authority, not inventing it. And links added where they don't help a reader would be noise, not signal. THE MOVE Add the internal links above — relevant, descriptive, genuinely useful — then leave the page alone and measure position and clicks in Search Console (Chapter 27) over weeks, not days. THE LESSON A great page nothing links to is a page in no race. Internal links put it on the track; the earlier fixes gave it legs.
Now step back and see what has happened to this page across four chapters, because this is the payoff the book has been building. Chapter 1's Figure 1.3 named the specific things the top result had and the #11 page didn't. Chapter 3 matched the intent. Chapter 9 fixed the title and made the page comprehensive. And this chapter added the internal links from pages that already have authority. The canonical promise from the very first pages of this book was precise: fix the title, add internal links from pages that already have authority, and cover the follow-up questions, and this page moves. Those three things are now done. The page is, at last, positioned to move.
Read that sentence carefully, because its honesty is the whole point. **"Positioned to move" is not "will rank
1," and we will never say otherwise.** We have removed the specific, fixable deficiencies that were holding a
genuinely good page down: it now matches intent, it is comprehensive, and it receives the internal authority its own site can give it. What remains is the one lever that depends on the outside world — external authority, the referring links from other sites that Chapters 22 through 24 are about — and that is a slower, longer game we do not control. So the honest claim is the strong one: we have made the page deserve a better position and given it every advantage its own architecture can provide. Google, weighing it against competitors we don't control, still makes the final call. We improved the odds as far as on-page and architecture can improve them, and that is exactly the job.
🛠️ Try It on Your Site Open Google Search Console (free) and go to the Links report, then Internal links. It lists your pages ranked by how many internal links point to each. Two things to look for: (1) your most important commercial pages should be near the top of that list — if a money page has only a handful of internal links, you have found a lever; (2) scroll to the bottom and find pages with one or zero internal links — those are underpowered, and some may be orphans (§15.7). Then pick your single most important underpowered page and add three genuinely useful internal links to it from your strongest, most relevant pages. That is the highest-ROI hour in technical SEO, and it costs nothing.
15.6 Hub-and-spoke and silos: organizing by topic
Individual internal links are tactics. Hub-and-spoke and silos are the patterns — the deliberate shapes you give a whole section of internal links so that a group of related pages reinforces itself. Both are ways of answering one question: how do you organize a body of related content so that Google understands it as a coherent topic, and so that authority concentrates within it rather than dissipating?
A hub-and-spoke structure is a topical arrangement in which one central "hub" page covers a subject broadly and links out to many focused "spoke" pages that each cover one sub-topic in depth — and every spoke links back to the hub (and often to its sibling spokes). The hub is the roundhouse; the spokes are the tracks radiating out. Here is the shape:
HUB-AND-SPOKE [schematic]
┌───────────────────────┐
│ HUB: "Water Heaters" │ ← broad pillar page; links to every spoke
│ (the plumbing hub) │
└───────────────────────┘
▲ ▲ ▲ ▲ ▲
┌────────────┘ │ │ │ └────────────┐
│ ┌────────┘ │ └────────┐ │
┌──────┴─────┐ ┌┴─────────┐ ┌─┴──────────┐ ┌┴───────┴──┐
│ Replacement│ │ Repair │ │ Tank vs. │ │ Signs it's │ ← spokes: each a focused page,
│ cost │ │ │ │ tankless │ │ failing │ each links BACK to the hub
└────────────┘ └──────────┘ └────────────┘ └────────────┘ and to relevant siblings
If that structure feels familiar, it should: hub-and-spoke is the architectural implementation of the pillar-and-cluster content model you met in Chapter 8. The pillar page and its cluster (a content-strategy idea — what to write) become, in the site's link graph, a hub and its spokes (an architecture idea — how to link it). The two are the same thing seen from two sides, and doing both is how a topic cluster actually earns its keep: the hub gathers authority and passes it to the spokes; the spokes cover the sub-topics comprehensively and pass authority (and topical relevance) back to the hub; and the whole cluster signals to Google that this site covers the subject thoroughly — the topical authority idea from Chapter 4.
🔗 Connection The pillar-and-cluster content model — deciding which pillar and which cluster pages to create — is Chapter 8 (Content Strategy), and the idea that covering a topic comprehensively builds topical authority is Chapter 4. This section is the architecture half: how to wire a cluster together with internal links so the content strategy actually pays off. Build the cluster in Chapter 8's terms; link it in this chapter's.
A silo is a related but broader idea: grouping a site's content into distinct thematic sections, and
keeping the internal links largely within each section, so that each theme is a self-reinforcing unit. A
home-services site might silo its content into Plumbing, HVAC, and Electrical, each with its own hub and
spokes, each section's pages linking mostly to other pages in the same section. Siloing can be physical (the
URL structure and folders reflect the themes — /services/plumbing/…) or virtual (the internal links create
the grouping regardless of URLs). The intent is to concentrate topical relevance and link equity within a
theme so Google sees a strong, coherent cluster rather than a scattered muddle.
Silos are useful, and they are also where two myths breed. Here is the first.
🚫 SEO Myth: "Use nofollow on internal links to 'sculpt' PageRank toward your money pages." This is a zombie tactic from the late 2000s. The theory: if you put a
nofollowattribute on the internal links you don't care about (login, cart, contact), the PageRank that would have flowed through them gets redirected to your remaining, "important" links — you "sculpt" the flow. It stopped working in 2009, when Google (via Matt Cutts) publicly changed hownofollowhandles PageRank: a nofollowed link no longer passes its share to the other links on the page — that share simply evaporates. So nofollow-sculpting doesn't concentrate PageRank; it throws some away. Do not nofollow your own internal links to sculpt flow. The legitimate way to send more equity to a page is the honest one from §15.5: link to it more, from stronger and more relevant pages, and link to genuinely trivial pages less. Steer with real links, not with attributes that only lose you authority.
And here is the second, subtler myth — the one that turns a good idea into a bad, user-hostile one.
🚫 SEO Myth: "Siloing means you must never link between silos." Rigid siloing — the doctrine that a plumbing page may never link to an HVAC page or you will "leak" or "dilute" your silo — is folklore that hurts both users and SEO. Google is thoroughly capable of understanding topical grouping without you building walls, and a homeowner reading about a failing water heater may genuinely benefit from a link to your furnace-maintenance guide. Relevance, not rigid separation, is the rule. Group related content and link it densely within a theme, yes — that is the sound core of siloing. But when a cross-topic link genuinely serves the reader, make it. An internal-linking strategy that forbids useful links to protect an abstract "silo" has forgotten that the entire point of the structure was to serve the reader in the first place. Theme one, again: the honest version of every technique is the one that helps the human; the manipulative version is the one that hurts them to please an imagined algorithm.
The takeaway across both patterns: hub-and-spoke and silos are ways of organizing genuine content by genuine topic so that structure amplifies it. What they can do is help Google see a coherent, comprehensive treatment of a subject and concentrate authority within it. What they cannot do is make thin content substantial or turn a link scheme into topical authority. Structure amplifies real coverage; it cannot fake it.
15.7 Orphan pages and crawl traps: the two failures that waste the most
Two architecture failures are worth isolating because they quietly cause more wasted effort than any subtle link-flow inefficiency. They are opposites: one is a page with too few links pointing to it, the other a structure that generates too many junk pages for the crawler to drown in.
An orphan page is a page that no other page on the site links to. It is stranded — an island. Because Google discovers pages mainly by following links (Chapter 1), an orphan is a page Google may struggle to find at all: it might sit in your XML sitemap and get discovered that way, or it might never be crawled, and even if it is found, it receives zero internal link equity because, by definition, nothing internal points at it. Orphan pages are shockingly common — they appear whenever a page is published but never linked from the navigation or any other page, when a section is redesigned and old pages are cut loose, or when a URL is changed and the internal links still point to the old address.
ORPHAN PAGE [schematic]
HOME ──▶ Services ──▶ Plumbing hub ──▶ Water heater page
│ │
└──▶ HVAC hub └──▶ Drain cleaning page
Panel Upgrades page ● ← nothing links here. In the sitemap only.
An ORPHAN: no internal links in, no equity,
discovered late or not at all.
Rivertown has exactly this problem, and you were told to hold the thought back in Chapter 1: three of its newer service pages — added by a since-departed contractor — are linked from nowhere. They are true orphans, which is precisely why Google "keeps taking weeks to find them" and why they rank for nothing. The fix is not mysterious. It is the same move as §15.5, applied for discovery instead of for lift: link to them — from the relevant service hub, from the navigation where appropriate, from related pages — so they join the site's link graph and start receiving both crawl attention and equity. An orphan is the easiest architecture problem to fix and one of the most common to have.
🔎 How Search Sees It Here is the cruel irony of an orphan page: it can be a genuinely excellent page and still get nothing, because quality is a ranking question and the page never reaches the ranking stage. Recall the pipeline — discover, crawl, render, index, rank. An orphan often fails at the very first stage, discovery. All the comprehensiveness and intent-matching in the world cannot help a page Google hasn't found, or has found but considers so unimportant (nothing links to it, so why would it be important?) that it rarely crawls it. This is why "why isn't my great new page ranking?" so often has an architecture answer, not a content one — and why an SEO's first question about an underperforming page is not "is it good?" but "what links to it?"
The opposite failure is a crawl trap: a structure that generates a near-infinite number of low-value URLs for Googlebot to crawl, wasting its attention on junk instead of your real pages. Crawl traps are usually created accidentally by systems that produce URLs on the fly. The classic culprits:
- Faceted navigation — filter-and-sort systems on category pages (
?color=red&size=large&sort=price) that multiply into thousands or millions of parameter-based URL combinations, each a slightly different page. This is the single biggest crawl trap in e-commerce, and it gets its full treatment in Chapter 31. - Endless calendars — a "next month" link that goes on forever, generating an infinite series of empty future dates.
- Session IDs or tracking parameters in URLs — the same page appearing at countless unique URLs because a parameter changes per visit.
- Internal search result pages — if crawlable and linked, these can generate unlimited thin URLs.
CRAWL TRAP [schematic]
Category page
├─ ?sort=price
├─ ?sort=price&color=red
├─ ?sort=price&color=red&size=large
├─ ?sort=price&color=red&size=large&page=2
└─ …thousands of near-identical combinations… ← Googlebot burns crawl budget HERE
instead of on your real pages
For most small sites — including Rivertown — crawl traps are not a live problem, and (as Chapter 1's myth-bust
about crawl budget noted) you should not lie awake over crawl budget on a few-hundred-page site. But you should
recognize the pattern, because the moment a site adds filters, a large catalog, or a calendar, the trap can
appear overnight. The tools that manage it — robots.txt to disallow crawling of parameter patterns, the
canonical tag to consolidate duplicates, and careful use of noindex — are all from Chapter 14; scaling
crawl control for genuinely large sites, with log-file analysis to see what Googlebot actually crawls, is
Chapter 33. Here the lesson is architectural awareness: know that infinite, low-value URL spaces are a
failure mode, and design so you don't accidentally build one.
🔄 Check Your Understanding Two pages on a site get no organic traffic. Page A is a detailed, well-written guide that no other page links to. Page B is one of ten thousand near-identical filtered URLs generated by the site's product filters. Which pipeline stage is failing for each, and what is the fix for each?
Answer
Page A is an orphan — it fails at discovery (and receives no internal equity). The fix is to link to it from relevant, ideally strong, pages so Google can find it and it gains authority. Page B is part of a crawl trap from faceted navigation — the failure is wasted crawling (and index bloat) on near-duplicate URLs. The fix is to stop Google from crawling/indexing the junk combinations — viarobots.txt, canonical tags, ornoindex(Chapter 14) — so crawl attention returns to the real pages. Opposite problems, opposite fixes: one page needs more links in; a class of pages needs to be kept out of the crawl.
📈 The Strategy File
It is time to give Rivertown Home Services something it has never had: a deliberate architecture. Recall the frozen facts. Rivertown is the family firm founded in 1984 by Ray Delgado and now run by his second-generation children, Marisa and Tony Delgado — HVAC, plumbing, and electrical, five locations (Rivertown HQ, Cedar Hills, Northgate, Westbrook, Millhaven), on WordPress, with about 8,000 mostly-branded organic visits a month. Its architecture today is not a design; it is an accident — a homepage, five thin location pages, a scattering of service pages (some orphaned), and a neglected blog, with almost no internal linking tying any of it together. We are going to give it a spine.
The design principle is the one this chapter has been building toward: a logical hierarchy that flows from the homepage down through service hubs to specific service-and-city pages, wired together with internal links so authority reaches the pages that must compete. Here is the shape.
FIGURE 15.2 — "Rivertown's designed architecture" [the Strategy File]
HOME (rivertownhome.example) depth 0
│ main nav points at the three service hubs + a locations hub
│
├──▶ PLUMBING HUB /services/plumbing/ depth 1 ← the hub we build first
│ ├──▶ Water Heater Replacement /services/plumbing/water-heater-replacement/ depth 2 ← the #11 page
│ ├──▶ Drain Cleaning /services/plumbing/drain-cleaning/
│ ├──▶ (the 3 orphaned service pages, now linked here — no longer orphans)
│ └──▶ each service → its 5 city pages:
│ /services/plumbing/water-heater-replacement/cedar-hills/ depth 3 ← service × city
│ …northgate/ …westbrook/ …millhaven/ …rivertown/
│
├──▶ HVAC HUB /services/hvac/ (same pattern: services → city pages)
├──▶ ELECTRICAL HUB /services/electrical/ (same pattern)
│
├──▶ LOCATIONS HUB /locations/ depth 1
│ └──▶ Cedar Hills / Northgate / Westbrook / Millhaven / Rivertown (each links to
│ the services offered there — cross-linking the two ways of slicing the same matrix)
│
└──▶ GUIDES / BLOG /guides/ depth 1
└──▶ "signs your water heater is failing," "tank vs. tankless," "how to flush a
water heater" → each links to the Water Heater Replacement page (feeding the #11 page)
BREADCRUMBS on every page: Home › Services › Plumbing › Water Heater Replacement › Cedar Hills
→ every deep page passes equity UP to its hub, reinforcing the hubs.
Walk through what this design does, decision by decision, because each one is a chapter concept made concrete.
It creates three topic hubs — Plumbing, HVAC, Electrical — each a hub-and-spoke cluster (§15.6) that matches the entity/topical territory mapped back in Chapter 4. The main navigation points at these hubs and the locations hub, and nothing else — a curated menu (§15.3), not a forty-item dump. This concentrates the homepage's authority on a handful of important destinations instead of scattering it.
It keeps important pages shallow. Every service is two clicks from home (home → hub → service); every service-and-city page is three (home → hub → service → city). No money page is buried in the basement (§15.1), and because the hubs are in the main nav, the crawler reaches everything quickly.
It uses clean, hierarchical URLs (§15.2) that mirror the structure: /services/plumbing/water-heater-
replacement/cedar-hills/. A human can read the address and know exactly what the page is and where it sits,
and Google gets the same hierarchy for free.
It fixes the orphans. The three stranded service pages from Chapter 1 now hang off the relevant hub — linked, discoverable, and receiving equity (§15.7). They rejoin the race.
It finishes the #11 page. The water-heater replacement page now receives internal links from the plumbing hub (its parent), from the homepage's guides module, and from three trafficked blog posts (§15.5, Figure 15.1) — the last of the fixable gaps Chapter 1 named. Combined with the Chapter 9 title and comprehensiveness work, the page is now positioned to move — which, as always, means we have earned it a fairer hearing, not promised it a ranking.
Now the discipline this book insists on: what this increment settles, and what it does not.
📄 Read the Report — the architecture increment, honestly bounded
text FIGURE 15.3 — "What the architecture design settles — and what it doesn't" [the Strategy File] THE PAGE Rivertown's new site architecture: hubs, shallow depth, clean URLs, breadcrumbs, internal linking, orphans fixed, the #11 page fed. WHAT IT SETTLES The site now has a spine: a logical hierarchy, curated navigation, hub-and-spoke topic clusters, important pages kept shallow, orphans linked, and — at last — the internal links that close the final fixable gap on the #11 page. Authority now has somewhere to flow and a path to the pages that must compete. WHAT IT DOESN'T The service×city pages are a SKELETON, not finished pages. Designing ~75 of them WITHOUT creating thin doorway spam is a serious problem in its own right — that is Chapter 33 (programmatic SEO), and we do NOT solve it here. The local-pack, Google Business Profile, and NAP work for five locations is Chapter 25. External authority is Chapters 22–24. And a good structure guarantees no ranking — it removes obstacles and steers earned authority; Google still decides. THE MOVE Build the three hubs and the URL/breadcrumb/nav structure now; link the orphans and feed the #11 page today. Leave the service×city pages as stubs to be built properly in Chapter 33, and the local layer for Chapter 25. THE LESSON Architecture is the frame the house is built on. It decides where authority CAN go; the content and links decide whether it's deserved. Frame first, then build.
What this adds to the running strategy. Rivertown now has an architecture blueprint — the third major component of the strategy file, joining the intent map (Chapter 3) and the on-page template (Chapter 9). When we reach the service×city system in Chapter 33, the local-SEO playbook in Chapter 25, and the full audit in Chapter 38, they all hang on this frame. (All Rivertown details are a constructed teaching example; the authority numbers in Figure 15.4's model are round illustrations, not measured values.)
Your Strategy-File task (on your own site, using Appendix C's architecture worksheet): sketch your site as a tree — home at the top, hubs beneath, everything else below that. Mark the crawl depth of your five most important pages. Circle any that are more than three clicks deep, and any page nothing links to (check the Search Console Internal Links report). Then draw the three internal links you will add this week to feed your most important underpowered page. Don't rebuild everything — make the frame visible and fix the two or three things that are obviously wrong. That is where architecture ROI lives.
Conclusion
We set out to explain why two equally good pages can have completely different fates, and the answer turned out to live in the space between pages — in site architecture, the organization and internal linking that most site owners never touch and that quietly decides which of their pages ever get a chance. We measured pages by crawl depth and learned to keep important ones shallow without worshipping a three-click rule. We designed URLs that are descriptive, hierarchical, and stable, while staying honest that a URL is a tiny ranking signal and a large human one. We used navigation, breadcrumbs, and footers as tools that serve the crawler and the reader at once. And we reached the lever the chapter was really about: internal links, which distribute link equity — the internal expression of PageRank — through a site, pooling authority where you point it and thinning where you spread it too far.
From there the strategy wrote itself: point your strongest pages at the ones that need to compete (§15.5); organize related content into hub-and-spoke clusters and silos that amplify genuine coverage (§15.6); and hunt down the two failures that waste the most — orphan pages that no link can reach and crawl traps that drown the crawler in junk (§15.7). We were careful about the boundary that keeps all of this honest: internal linking redistributes the authority you have earned; it does not manufacture it, and no clever internal pattern substitutes for the external links of Chapters 22–24. That boundary is theme five in practice, sitting on the technical foundation of theme four.
And we closed a thread the book opened on its first pages. The page stuck at #11 — Rivertown's water-heater guide — has now had its title fixed (Chapter 9), its comprehensiveness built (Chapter 9), and, here, the internal links added from pages that already have authority. The three fixable gaps Chapter 1 named are closed. The page is positioned to move. Not promised a position — positioned, which is the honest and the better claim: we made it deserve better and gave it every advantage its own site can provide, and left the rest, correctly, to earned authority and Google's judgment.
A site can be perfectly structured and still lose if it is painful to use — slow to load, jumpy on a phone, frustrating to interact with. Google measures that experience directly now, and it is a ranking input in its own right. That is where we go next.
→ Continue to Chapter 16: Core Web Vitals.
Key Terms
- Site architecture — the overall organization and internal-link structure of a website: how pages are grouped, nested, and connected, from URL hierarchy to navigation to internal linking.
- Flat architecture — a structure in which most important pages are reachable in a small number of clicks from the homepage (the site is "wide"); generally good for discovery and link flow, up to the point where excessive width dilutes signal.
- Deep architecture — a structure in which reaching content requires many clicks down through nested layers (the site is "tall"); risks leaving important pages crawled rarely and starved of authority.
- Crawl depth — the number of link "hops" (clicks) it takes to reach a page from the homepage; the homepage is depth 0. Also called click depth. Matters more than the number of slashes in the URL.
- URL structure — the deliberate design of a site's URLs: descriptive words, a hierarchy of folders that mirrors the site, and stable addresses that don't change.
- Breadcrumb — a secondary navigation trail showing a page's place in the hierarchy (e.g.,
Home › Services › Plumbing); aids user orientation, adds internal links up the hierarchy, and passes equity to parent pages. - Internal link — a link from one page to another page on the same site; helps Google discover and understand pages, and distributes authority among them.
- Link equity — the ranking value a link passes from the source page to the target page; the internal expression of PageRank (informally "link juice"). It flows through internal links and thins with distance.
- Orphan page — a page that no other page on the site links to; hard for Google to discover and receiving zero internal link equity, so it may never reach the ranking stage no matter how good it is.
- Hub-and-spoke — a topical structure in which a central "hub" page covers a subject broadly and links to focused "spoke" pages that each cover a sub-topic and link back; the architectural implementation of the pillar-and-cluster content model (Chapter 8).
- Silo — the grouping of a site's content into distinct thematic sections whose internal links stay largely within each section, concentrating topical relevance and link equity within a theme (physical via URLs/folders, or virtual via links).
Spaced Review
Retrieval practice mixing this chapter with earlier ones. Try each before revealing the answer.
- Explain the difference between crawl depth and URL path depth, and which one Google's guidance says matters more for how important a page is treated.
- What is link equity, how does it flow through a site, and why does a link from a page with few outgoing links pass more of it than a link from a page with hundreds?
- (From Chapter 1.) Chapter 1's Figure 1.3 named the specific gaps holding the water-heater page at #11. Which gap did this chapter close, and why is "the page is now positioned to move" a more honest claim than "the page will now rank #1"?
- (From Chapter 14.) A site has ten thousand near-identical filtered URLs from its product facets clogging the crawl. Name two Chapter 14 tools you would use to keep Google from crawling or indexing the junk — and say why this is a crawl-trap problem, not an orphan problem.
- (From Chapter 1.) Google discovers pages mainly by following links. Given that, explain in one sentence why an excellent page that nothing links to can still get zero traffic — and name the pipeline stage that fails first.