The Scavenger and the Blueprint: Two Paths to a Readable Web
When we talk about making public information truly public, we tend to celebrate two distinct virtues: completeness and clarity. The archivist wants to save everything, ensuring no record is lost. The communicator wants to structure everything, ensuring no record is misunderstood. In the messy world of open data and public records, these virtues often manifest as two contrasting approaches, which I’ve come to think of as the Scavenger and the Blueprint.
The Scavenger approach is one of salvage and assembly. It accepts the digital landscape as it is—fragmented, inconsistent, and sprawling—and seeks to gather what exists, wherever it exists. This is the work of the web archivist capturing a city council’s disparate PDF minutes from a dozen different subdomain URLs, or the researcher stitching together a narrative from social media posts, cached news articles, and scanned meeting handouts. The Scavenger’s primary tool is often a browser bookmark and a keen eye for digital ephemera. Their victory is in the recovery itself, in proving that a record existed at all. The resulting archive is rich, often surprising, but it can also be a cabinet of curiosities—a collection of artifacts whose relationships and contexts are implied, not explicitly defined.
The Order of the Blueprint
In stark contrast stands the Blueprint approach. This method is not about foraging in the existing digital undergrowth, but about designing the ground from which information will grow. It is the work of the data standardist, the semantic web developer, the civic technologist drafting an open records schema. The Blueprint asks not “What can I find?” but “How should this be built so it can always be found and read?” Its goal is to impose a prior order, creating templates and protocols so that data is born structured, linked, and machine-readable. The victory here is in foresight and interoperability, in creating systems where information slots neatly into place.
Both are essential, yet they often view each other with a quiet skepticism. To the Scavenger, the Blueprint can feel like a beautiful, sterile theory—a plan for a city that will never be built, while real historical fragments are blowing away in the digital wind. They wonder: what good is a perfect schema for building permits if the actual permits from the last decade are locked in a proprietary format on a defunct server? To the Blueprint builder, the Scavenger’s work can seem like a desperate, endless reaction—a perpetual game of catch-up that does nothing to stop the next wave of loss. They ask: why spend a thousand hours rescuing badly formatted data when we could spend a hundred hours ensuring the next batch is born free?
The truth is, our readable public record needs both. We need the Scavengers to rescue the past and document the chaotic present, proving through sheer accumulation what is at stake. Their work provides the urgent, human case for preservation. And we need the Blueprint makers to architect a more sensible future, to slowly bend institutional practice toward sustainable openness. One approach answers the emergency; the other tries to rewrite the fire code.
In the end, the most durable digital heritage is likely built by those who understand both mindsets: who can salvage a critical dataset with one hand while drafting a standard with the other, knowing that preservation is always both an act of recovery and an act of hope for a less fragmented tomorrow.
Notes & further reading
A few pages I came back to while writing this:
- Washington, DC
- The Digital Dustpan: On Sweeping Up Our Own Ephemera
- one area's overview
- The Vernal Equinox of the Public Record: On Data, Daylight, and Balance
- a practical rundown
- The Silent Archive: On the Politics of What We Choose Not to Save
- Little Rock, AR
- Gilbert, AZ
- Peoria, AZ
- Surprise, AZ
- Elk Grove, CA
- Pasadena, CA
- New Haven, CT