The earliest pages of the web are disappearing faster than most people realize. Broken links, shuttered hosting providers, and abandoned domains have quietly erased a significant share of content from the internet’s first decade.
The scale of the loss
Researchers estimate that a large share of links from the late 1990s and early 2000s no longer resolve to their original content. Personal sites, early forums, and independent publications have been especially vulnerable, since they often lacked the institutional backing to survive server migrations or ownership changes.
Every year we wait, more of the early web simply stops existing. There’s no second copy once the server goes dark.
Archivists step in
Organizations dedicated to web preservation have accelerated efforts to crawl and store as much of the surviving early internet as possible, race against expiring domains and aging hardware that could fail at any time.
Why it matters beyond nostalgia
The first decade of the web captured how ordinary people, not just institutions, experimented with publishing, community, and communication online. Losing that record erases a foundational chapter of internet and technology history that later platforms were built on. Preserving it now, while fragments still exist, is far easier than trying to reconstruct it later.