/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

← → days · ↑ ↓ browse · Enter similar · o open

Hachette, Penguin Random House, Wiley, HarperCollins sue Internet Archive, saying its project to let users borrow ebooks scanned from books violates copyrights

Elizabeth A. Harris / New York Times :

New York Times Elizabeth A. Harris

Context & Ripple Effects

This June 2020 suit is the opening move in what became the defining copyright fight of the decade for books. Four major publishers — [[a:none|Hachette]], Penguin Random House, Wiley, and HarperCollins — argued that scanning physical books and lending them one-at-a-time online was infringement, not library service. The case turned on a deceptively simple question: who owns an ebook, and does buying a paper copy confer any right to lend its digital twin?

The arc that followed validated the publishers' strategy. A US judge sided with them in March 2023, and the Internet Archive's subsequent appeal kept the fight alive while the Archive agreed to drop the publishers' full book catalogs from its lending program. The same plaintiff coalition then carried the doctrine into a new arena, suing Meta and Google over AI models trained on copyrighted books.

First-order effects

  • The Internet Archive faces an existential legal threat to its Open Library: if courts accept the publishers' framing, every scanned-and-lent title is infringing regardless of how few 'copies' circulate at once.
  • The four publishers establish a test case they control — small enough defendant, clear-cut conduct — that lets them set the boundary of digital lending without touching their own licensing businesses.

Second-order effects

  • A publisher win forces other digital-lending operations — library consortia, school platforms — to license rather than scan, channeling revenue through publisher-controlled terms and raising costs for underfunded libraries.
  • The legal theory proven here becomes reusable ammunition: the same publishers, joined by Scott Turow, later deploy it against far richer targets in class-action suits against Meta and Google over Gemini training data.

Third-order effects

  • If the pattern holds, copyright enforcement consolidates around litigation-first coalitions of large publishers and authors' representatives, with nonprofit archives and AI developers alike treated as unlicensed distributors.
  • The case pushes the industry toward a regime where no copy exists outside a license — physical ownership grants nothing digitally — concentrating pricing power over ebooks, library lending, and eventually AI training corpora in publisher hands.

The trend: Book publishing is extending copyright enforcement from pirate sites to any unauthorized digital use — scanned lending first, AI training next — using the same plaintiff playbook across each new copying technology.

Discussion

  • @ornithophobix Shalmi on x
    Shame on these big-name publishers for trying to restrict fair use of books when the pandemic has effectively made it impossible to access physical libraries. @internetarchive is an invaluable scholarly resource and must be protected at all costs. https://www.theverge.com/...
  • @digitaldutta Srinivas Kodali on x
    The @internetarchive is being sued by publishers and this is going to be an important copyright case. https://publishers.org/...
  • @nytimesbooks @nytimesbooks on x
    Penguin Random House, HarperCollins, Hachette and Wiley accused Internet Archive of piracy, for making over 1 million books free online during the coronavirus outbreak https://www.nytimes.com/...
  • @liz_a_harris Elizabeth A. Harris on x
    Internet Archive has made more than a million books available for free online in what it called an effort to “serve the nation's displaced learners.” Publishers and authors call it something else: theft. https://www.nytimes.com/...
  • @nytimestech @nytimestech on x
    A group of publishers sued Internet Archive, saying that the nonprofit group's trove of free electronic copies of books is robbing authors and publishers of revenue at a moment when it is desperately needed https://www.nytimes.com/...