Chapter 33 Self-Check Quiz

24 questions — multiple choice and short answer. Answers are in the collapsed key at the bottom. Try the whole set before you reveal it.

Multiple choice

1. Programmatic SEO is best described as: - A) A black-hat technique for tricking Google with automated pages - B) Creating many pages at scale from a template fed by a structured dataset, one page per record - C) Using artificial intelligence to write blog posts - D) Buying links to thousands of pages at once

2. Which statement about how Google treats a templated page is correct? - A) Google penalizes any page it detects was generated from a template - B) Google gives templated pages a small ranking boost for consistency - C) Google evaluates the finished page's value and does not care whether it was templated or hand-written - D) Google cannot crawl template-generated pages

3. A doorway page is: - A) The homepage of a large site - B) A page created mainly to rank for a query or location while offering no real value of its own, funneling users onward - C) Any page with a contact form - D) A page blocked by robots.txt

4. The difference between a legitimate service×city page and a doorway page is primarily: - A) Word count - B) How many keywords appear in the title - C) Whether the page contains real, differentiated local content rather than spun text - D) Whether it uses schema markup

5. Google names two forces that shape crawl budget. They are: - A) Crawl capacity limit and crawl demand - B) PageRank and relevance - C) Sitemap size and robots.txt length - D) Impressions and clicks

6. The crawl capacity limit goes up when: - A) You add more pages to the site - B) Your server responds quickly and without errors - C) You submit a larger sitemap - D) You buy more backlinks

7. Which site most plausibly needs to actively optimize crawl budget? - A) A 40-page local plumber's site - B) A 75-page service×city site like Rivertown's - C) A retailer with 2 million URLs and heavy faceted navigation - D) A five-page brochure site

8. Log-file analysis primarily tells you: - A) Why a page ranks where it does - B) What Googlebot actually crawls, how often, and where crawl is wasted - C) Your conversion rate - D) Which keywords to target

9. Before trusting "Googlebot" lines in your server logs, you should: - A) Assume they are all genuine - B) Verify them with a reverse/forward DNS lookup (or Google's published IP ranges) - C) Block them in robots.txt - D) Ignore any request over 10,000 bytes

10. In a templated site, the danger of a single template change is that: - A) It only affects one page, so it's not worth doing - B) A mistake (bad canonical, accidental noindex, broken schema) ships to every page built from the template at once - C) Google will re-crawl the whole web - D) It cannot be rolled back

11. "More pages means more traffic" is a myth because, past the point of real demand, extra thin pages mainly cause: - A) Faster crawling - B) Crawl waste, index bloat, and a lower site-wide quality signal - C) Higher click-through rate - D) More backlinks

12. In the enterprise setting, the binding constraint on SEO is usually: - A) Not knowing what to do technically - B) Coordination — getting changes prioritized, built, approved, and shipped across teams - C) The cost of keyword tools - D) Google's crawl budget

13. An SEO ticket is most likely to actually ship when it includes: - A) Only the phrase "improve our SEO" - B) A clear change, the business why, the scope, acceptance criteria, and a priority - C) A demand that engineering drop all feature work - D) A long list of every possible fix

14. Which internal-linking pattern is a spam risk at scale? - A) Each leaf page links up to its service hub and city hub - B) Nearby-city and same-city sibling links - C) Thousands of automated internal links all using the identical exact-match anchor text - D) A breadcrumb trail on each page

15. To remove 10,000 thin parameter URLs from the index, the correct first move is: - A) Disallow them in robots.txt immediately - B) Allow crawling and apply noindex so Google can read the removal instruction, then consider robots.txt after they drop out - C) Delete the entire site section - D) 301-redirect all of them to the homepage

16. Rivertown's ~75 service×city pages, built well, will primarily win: - A) The local pack (map pack) in all five cities on their own - B) Localized organic rankings for "{service} {city}" queries, while supporting the Business Profiles - C) Nothing, because location pages never rank - D) Paid-search auctions

17. "Rule 0 — subtract first" in the service×city system means: - A) Always build every cell in the matrix - B) Build only the cells with real demand and a real local story; skip or merge the rest - C) Delete the homepage - D) Remove all internal links

18. A store with 12,000 products shows 600,000 indexed URLs. The most likely programmatic cause is: - A) Too few backlinks - B) Unmanaged faceted navigation generating parameter URLs at scale (Chapter 31) - C) A slow server - D) Missing meta descriptions

Short answer

19. State, in one sentence, the single thing that distinguishes a legitimate location page from a doorway page.

20. Explain why "the method of production is invisible to Google, but the value of the result is not" is the key idea of this chapter.

21. Give two concrete sources of crawl waste that a log-file analysis commonly uncovers, and the fix for each.

22. Why is a beautifully written service×city page still at risk of never ranking if it is an orphan, and what internal-link structure prevents that?

23. Describe the two-layer structure of a good location page and say which layer earns the page its right to exist.

24. What is governance in enterprise SEO, and what does it protect a programmatic system from over time?


Answer key 1. **B** — template + structured data → many pages, one per record. 2. **C** — Google evaluates the finished page's value; templating is invisible to it. 3. **B** — the defining features are "no real value of its own" and "funnels users onward." 4. **C** — real, differentiated local content, not word count, keywords, or schema. 5. **A** — crawl capacity limit and crawl demand. 6. **B** — a fast, healthy server raises the capacity ceiling; Google throttles to server health. 7. **C** — crawl budget matters at genuine scale (millions of URLs, heavy facets). The small sites do not need it. 8. **B** — logs are about crawling, not ranking. 9. **B** — reverse/forward DNS (or published IP ranges); spammers spoof the user-agent. 10. **B** — leverage runs both ways; a template bug is a site-section bug. 11. **B** — crawl waste, index bloat, and a diluted site-wide quality signal. 12. **B** — enterprise SEO is a coordination and governance problem more than a knowledge problem. 13. **B** — what, why (impact), scope, acceptance criteria, priority. 14. **C** — identical exact-match anchors across thousands of automated links is an engineered, manipulative pattern (Chapter 24). The others are healthy. 15. **B** — `noindex` first (Google must crawl to read it), `robots.txt` only after de-indexing. Redirecting all to the homepage creates soft 404s. 16. **B** — the pages win localized organic rankings and support the profiles; the local pack is a separate build (Chapter 25). 17. **B** — subtract first: build the cells with real demand and a real local story. 18. **B** — unmanaged faceted navigation (parameter URLs) is the classic index-bloat cause. 19. Whether the page contains **real, differentiated local content** specific to that service in that place — something a find-and-replace could not produce — rather than spun/duplicated text. 20. Because there is no penalty for *how* a page is made (templated, AI-assisted, hand-typed) — Google crawls and judges the finished result. So you cannot be punished for using a template, and you cannot be saved by disguising a valueless page as hand-made. All the effort should go into the value of the result, which is what Google actually measures. 21. Any two, e.g.: **parameter/faceted URLs** → canonicalize sort/view variants and promote high-demand facets to real pages, then `noindex`/`robots.txt` the rest; **redirect chains** → flatten to a single hop and update internal links to the final URL; **soft 404s / real 404s** → return honest status codes and fix the broken links; **duplicate content** → consolidate with `rel="canonical"`. 22. Because internal links are how Google discovers pages and how link equity flows (Chapter 15); an orphan — linked from nowhere — most often fails at the *discovery* stage and gets zero internal authority, so quality alone never enters it in the ranking contest. The fix: the double hub-and-spoke — each leaf linked up to both its service hub and its city hub and across to relevant siblings, so nothing is orphaned. 23. A **shared service layer** (what the service involves, repair-vs-replace, safety, FAQ) that may be legitimately templated across cities, and a **unique local layer** (the local branch and team, the city's building/climate reality, local reviews and proof, permits) that must be real and differentiated. The *unique local layer* is what earns the page its right to exist. 24. Governance is the set of standards and gates — who may publish programmatic pages, who approves a template change, what quality bar a new batch must clear before launch, the schema/internal-linking standards — that keeps a large, multi-team, page-generating system from drifting into thin pages, broken templates, and index bloat over time. It protects the system from quietly rotting into a doorway farm as teams and priorities churn.