Posted on

Synthetic historical photographs circulating as ordinary archival images

Historical photographs feel different from illustrations because they appear to contain a physical trace of an event.

A camera was somewhere. Light hit film or a sensor. Somebody pressed a shutter. Even when a photograph is staged, cropped, selectively captioned, or poorly interpreted, there is usually an original object with a date, creator, negative, publication history, or archival chain that researchers can investigate.

A generated “historical photograph” can imitate the surface of all of that without any of it existing.

The problem becomes serious after the image leaves the generator.

A caption can manufacture an archive

In 2024, AAP FactCheck investigated a black-and-white Facebook image claiming to show Henry Ford sitting in his first automobile in 1896. The image looked unusually crisp and historical enough to attract attention, but it was AI-generated and did not accurately depict Ford or his quadricycle. AAP found it was one of multiple synthetic “historical” pictures being circulated by the same kind of page. See AI used to generate fake history photos.

The dangerous step is not merely generating the picture. It is attaching a confident caption.

“Henry Ford, 1896” changes an illustration into an apparent record. Repost it a few times, crop away the original context, and the image can begin appearing in collections, blogs, slideshows, family-history pages, and search results as if it came from an archive.

At that point visual inspection is a weak defense. Film grain, scratches, period clothing, lens softness, and damaged borders can all be synthesized too.

Provenance matters more than vibes

For a supposed archival photograph, useful questions are boring and powerful: Which archive holds it? Is there a catalog number? Who is the photographer? When was the image first published? Is there a negative, contact sheet, newspaper reproduction, or accession record? Can another institution independently identify it?

That chain is provenance.

The absence of provenance does not automatically prove an image is fake. Plenty of genuine family photographs survive with almost no metadata. But the less context an image has, the weaker the historical claim should become.

Synthetic imagery makes this standard more important because “looks old” no longer implies “was created in the past.”

The internet has always mislabeled photographs. Generative systems add a new category: photographs of events that never passed through a camera at all.

History needs more than sepia.

Posted on

Ancient Web: A Columbine Site Preserves an Important Warning About Fake Web Mirrors

A bad archive can preserve misinformation just as efficiently as it preserves files.

That is what makes this page from A Columbine Site worth reading carefully.

Visit the Eric Harris webpages archive page

A Columbine Site describes itself as a long-running research archive dedicated to the injured, the survivors and those who died in the April 20, 1999 Columbine High School shooting. Its Eric Harris webpage section deals with material that existed online before the attack and what happened to that material afterward.

The important part is not the killer’s writing. It is the site’s unusually explicit discussion of provenance.

The page says Harris had used several online services and maintained webpages, but that his actual pages were removed before their addresses became widely public. The site’s author says some directory material and files were preserved, while the original HTML for the webpages was already gone.

That created a problem.

A reconstruction escaped its label

The archivist explains that in 1999 they displayed surviving graphics and links in a layout intended to help readers imagine how some of the missing pages may have looked. Other websites later copied those displays and presented them as mirrors of Harris’s original webpages.

According to A Columbine Site, they were not.

The page now goes out of its way to distinguish actual preserved directories from later reconstruction pages and says there are no surviving copies of Harris’s original websites. Whether a reader is researching Columbine specifically or digital history generally, that distinction is valuable.

Once a screenshot, mirror or copied page loses its chain of custody, repetition can turn a reconstruction into supposed evidence. Search engines then index the copies. New writers cite the search results. Eventually the false provenance becomes easier to find than the correction.

This site is also a reminder that disturbing historical material needs context. A Columbine Site’s homepage frames its collection around victims, survivors, official documents, reporting and the consequences of the attack. That is a far more useful archival purpose than turning the perpetrators into Internet folklore.

For researchers in 2026, the lesson extends well beyond this event: “I found it online” is not provenance. A mirror may be genuine, partial, reconstructed, mislabeled or copied from another reconstruction. The URL alone cannot tell you which.

In this case, the surviving archive does something responsible and surprisingly rare.

It tells you exactly where its evidence stops.

Read the original archival explanation