Eight daily newspapers owned by Alden sue OpenAI and Microsoft, accusing them of using copyrighted articles without permission to train generative AI products
The suit, which accuses the tech companies of copyright infringement, adds to the fight over the online data used to power artificial intelligence.
Context & Ripple Effects
Alden’s papers join an escalating publisher challenge to AI training practices after the New York Times brought its copyright case against OpenAI and Microsoft and other digital publishers filed related claims. The dispute turns a broader concern over journalism’s reuse in generative systems into another direct legal test.
The case also follows signs that publishers were seeking compensation for AI use, including Dow Jones’s assertion that OpenAI had no agreement to use Wall Street Journal reporting. It matters because a large newspaper operator expands the set of news owners pressing the issue.
First-order effects
- OpenAI and Microsoft face another copyright suit over the material used to train generative AI products, while Alden’s newspapers seek to assert control over their articles’ use.
- The filing adds Alden’s titles to the group of publishers pursuing a legal, rather than solely commercial, response to unlicensed AI training.
Second-order effects
- The additional case raises pressure on AI companies to defend a consistent legal position across publisher claims and on publishers to weigh litigation against licensing discussions.
- It strengthens the relevance of the earlier copyright claims filed by Raw Story, AlterNet, and The Intercept, making news content provenance a more consequential issue for generative-AI providers and their partners.
Third-order effects
- If publisher claims continue to accumulate, court outcomes or settlements could help determine whether news archives become a licensable input to AI development rather than content providers can assume is available for training.
- The pattern points to a more formal market for rights, permissions, and attribution around high-value text data, though the governing copyright boundaries remain unsettled.
The trend: Generative-AI companies are confronting a shift from broadly available web data toward negotiated and litigated access to professionally produced content.