/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

← → days · ↑ ↓ browse · Enter similar · o open

The Atlantic, Politico, Vox and others sue Cohere, alleging it used copyrighted works to train its LLM and shared versions of entire articles without permission

Lawsuit accuses Canadian company of sharing versions of entire articles without permission  —  The Atlantic, Politico …

Wall Street Journal Alexandra Bruell

Context & Ripple Effects

This case places a Canadian LLM developer in the same copyright conflict already seen in Canadian news publishers' case against OpenAI. The allegations cover both model training and the sharing of article-length material, making the dispute about downstream product behavior as well as data acquisition.

The related coverage also shows the dispute widening beyond news: book and academic publishers have brought similar claims against major AI developers, including the publisher-led case against Meta.

First-order effects

  • Cohere must defend against claims from major publishers over its training data and alleged article-length outputs, creating immediate legal and product-risk exposure.
  • The publishers seek to establish that unauthorized use of their reporting in LLM development and distribution is actionable, rather than merely a licensing disagreement.

Second-order effects

  • Publishers and platforms may further tighten access to their material; Politico is already reported to be considering restrictions on Google's access for AI use.
  • Other model providers face added pressure to document content provenance, negotiate licenses, and limit outputs that could substitute for source articles.

Third-order effects

  • If courts treat training and article-like outputs as distinct copyright risks, AI companies may need separate controls for dataset sourcing and retrieval or generation behavior.
  • The pattern points toward content owners treating high-quality archives as commercial AI inputs, with litigation and licensing jointly shaping access terms.

The trend: Copyright disputes are moving from a narrow fight over web-scale training data toward a broader market for licensed content and constraints on how models reproduce it.

Discussion

  • @ednewtonrex Ed Newton-Rex on x
    Another generative AI lawsuit lands. 14 news publishers, including Condé Nast, The Atlantic, Forbes, The Guardian, Business Insider, the LA Times, Politico, Vox Media & many more, have sued @cohere, alleging “massive, systematic copyright infringement”. “The complaint provides [i…
  • @danprimack Dan Primack on x
    Cohere is prepping an IPO. Wonder if this puts any crimp in their plans