Posted on

Government website transitions and missing historical publications

Government websites change for ordinary reasons: a new administration arrives, an agency reorganizes, a contractor replaces the CMS, or an office decides that twelve years of nested directories would look better as one enormous search box.

The historical consequence can be less ordinary. Reports that were once directly linked may move to a new repository, receive new filenames, lose old landing pages, or disappear from navigation entirely. From the outside, relocation and deletion can look exactly the same: yesterday’s URL returns 404.

That problem is serious enough that preservation institutions created the End of Term Web Archive. The project began in 2008 and has collected U.S. federal websites around the 2008, 2012, 2016, 2020, and 2024 presidential transitions. The archive exists because an administrative transition is also a web transition.

A missing URL is only the first clue

When a government publication vanishes from its old address, the first task is not to declare it erased. Agencies often move documents between content systems, publications databases, records repositories, and redesigned websites.

Useful checks include searching the exact title, report number, filename, quoted phrases, agency publication catalog, and the current domain. A PDF may survive under a completely different path. An agency may have moved older material into an archival section without keeping redirects from the original URLs.

The White House provides an unusually clear example because the domain itself passes to the incoming administration. NARA’s archived presidential websites page preserves successive administrations separately and warns that archived sites are frozen historical records whose broken internal or external links are not repaired. A publication can therefore survive in the presidential archive even though the current whitehouse.gov no longer exposes it at the old address.

Archives help answer what changed

The End of Term collections are especially useful because they capture websites before and during known periods of change. They include HTML, images, PDFs, spreadsheets, multimedia, and other files stored in web-archive formats. NARA’s broader web-records guidance likewise treats long-term preservation of federal web content as part of maintaining the historical record.

That still does not prove every missing file was deliberately removed. Web crawlers miss things. Some databases require forms or scripts that are difficult to capture. Files may have been outside the crawl scope, blocked, generated dynamically, or linked only from pages the crawler never reached.

This distinction matters. “The current website no longer exposes this document at its old URL” is an observable fact. “The government deleted the document from existence” is a much stronger claim and often a false one.

Government web transitions create real historical gaps, but careful preservation lets us separate redesign damage from actual disappearance. Without that comparison layer, a routine CMS migration can look like a purge, while a genuine loss can hide behind the bland language of a website refresh.