Inside the world of digital hoarders, who enjoy collecting terabytes worth of data and take pride in archiving files that often disappear from the internet
This week, we are writing about waste and trash, examining the junk that dominates our lives, and digging through garbage for treasure.
Context & Ripple Effects
This Gizmodo piece sits inside a week-long series on waste and trash, profiling hobbyists who archive terabytes of files that platforms let disappear — an amateur answer to a problem the related coverage keeps documenting. The Internet Archive's growth from 2TB in 1997 to roughly 100PB shows institutional preservation scaling up, while BuzzFeed's earlier reporting argued tech firms and governments can't be trusted to preserve digital history on their own, leaving room for exactly these volunteer collectors.
First-order effects
- Hoarders' private drives become de facto backups for content that vanishes from the live web, complementing formal projects like the Internet Archive rather than duplicating them.
Second-order effects
- Cheap consumer storage turns hoarding into a low-cost hobby even as corporations take the opposite path — the FT investigation found Amazon and Microsoft destroy millions of drives yearly despite software wiping, discarding capacity individual archivers would happily reuse.
Third-order effects
- As Brewster Kahle's copyright fights with labels like UMG show, preservation increasingly collides with rights holders — and if that pressure constrains institutional archives, distributed personal collections become the resilient fallback layer. Meanwhile, a generation raised on search rather than files (teachers report students unfamiliar with directories and folders) may lack the file-management habits hoarding depends on.
The trend: Preservation of the disappearing web is shifting from institutions alone toward a hybrid of scaled archives like the Internet Archive and distributed amateur collectors.