The Unclicked Link: On the Fallacy of the Perfect Page Graph
There’s a prevailing faith in the architecture of the web that borders on the theological. We are taught that a well-structured site is a logical, interconnected web, a perfect graph where every node is reachable from another through a sensible path of links. The creed of internal linking is sacrosanct: every page should be just a few clicks from the homepage, and silos should be meticulously bridged. The goal is to present a flawless, crawlable map to the search engine’s spider, lest any precious content become a digital ghost. We speak of ‘PageRank’ and ‘link equity’ as if they were hydrological systems, flowing predictably from a central source through engineered canals to every corner of the domain.
But the web is not a plumbing system, and the crawler is not water. It’s an agent of chaos with a terrible memory, a creature that discovers not just by following a preordained path, but by getting lost. Our obsession with creating perfect, frictionless routes for discovery ignores a more fundamental truth: the journey is often more revealing than the destination. By making every path so explicit, so logical, we may be sanitizing the very clues that signal a page's true relevance and unique context.
Think of a library where every book is connected to every other by a literal string. A string leads from a history of Rome to a biography of Caesar, and from there to a text on ancient engineering. It’s efficient, certainly. But it eliminates the beautiful, serendipitous discovery of finding a book on Roman aqueducts mistakenly shelved in the poetry section, creating an unexpected connection between structure and art that the master cataloguer never intended. The misplaced book is an anomaly, and anomalies are memorable. They create stronger associative bonds precisely because they break the pattern.
An unclicked link—a piece of content with no internal inbound links—is often viewed as a failure of site architecture. We call it an ‘orphan page’ and fret over its isolation. Yet, this isolation can be its greatest strength. How does a crawler find it? Perhaps through an old, unlinked URL in a forum post from a decade ago. Perhaps through an external sitemap it consults separately from the main navigation. When it arrives, it finds a page that exists outside the main narrative of the site. It hasn’t been pre-digested by the site’s own internal logic. The crawler must assess it on its own terms, in its raw state, which can sometimes lead to a clearer, less biased understanding of its singular purpose.
This isn’t an argument for messy sites, but rather a challenge to the dogma of hyper-logical ones. The goal shouldn’t be to eliminate all dead ends and unconnected nodes, but to allow for a certain productive wilderness within your domain. A page that is reached by a single, obscure, external path carries a different kind of signal—one of specific intent, of a destination sought rather than a waypoint passed through. In our zeal to build the perfect graph, we risk building a site that is perfectly crawlable but utterly soulless, a place where every discovery is an expectation fulfilled, and no discovery is a surprise.
Notes & further reading
A few pages I came back to while writing this:
- Anchorage, AK
- The Indexer's Ghost: On the Unseen Hands that Built the First Map
- Birmingham, AL
- The Shoebox Under the Bed: On the Private Geography of the Unsubmitted Sitemap
- Huntsville, AL
- The Quiet Threshold: On the Unseen Labor of the Robots.txt Greeting
- Montgomery, AL
- Little Rock, AR
- Chandler, AZ
- Gilbert, AZ
- Mesa, AZ
- Peoria, AZ
- Phoenix, AZ