/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

← → days · ↑ ↓ browse · Enter similar · o open

Internal OpenAI docs detail contractors evaluating anonymized prompts and chats to improve the models; model training is turned on by default for consumer plans

404 Media Joseph Cox

Context & Ripple Effects

OpenAI drew a boundary in 2023 when it said API-submitted data would not be used for service improvements without customer opt-in. The consumer-plan policy described here makes that boundary commercially significant: the default treatment of consumer interactions differs from the earlier API opt-in commitment.

The disclosure also follows reports that OpenAI asked contractors to supply work from present or past jobs for evaluation while leaving them to remove confidential material. That contractor-supplied work review practice put responsibility for sensitive-data screening close to the people providing the material.

First-order effects

  • Consumer ChatGPT users’ prompts and chats can enter OpenAI’s model-improvement workflow by default, with contractors evaluating anonymized material.
  • OpenAI gains a human-review channel for assessing model behavior while placing greater weight on consumers recognizing and changing their training settings.

Second-order effects

  • OpenAI’s distinction between consumer defaults and its API opt-in policy gives business buyers a clearer reason to treat enterprise data controls as a separate procurement requirement.
  • Anthropic’s confirmation, cited in public discussion, that it also uses human reviewers makes reviewer access and disclosure practices a competitive trust issue across consumer AI services.

Third-order effects

  • If consumer services continue to rely on default-on data use and contractor review, AI providers will increasingly compete on how clearly they separate consumer learning loops from business-data protections.
  • The pattern points toward frontier-model governance extending beyond model outputs to the handling, review, and consent rules around the inputs used to improve them.

The trend: Consumer AI is turning everyday use into a model-improvement pipeline, making data defaults and human-review controls central to platform trust.

Discussion

  • @josephfcox Joseph Cox on x
    New from 404 Media: humans are reading ChatGPT conversations OpenAI has hired an army of contractors who read real ChatGPT users' chats. I've seen internal docs, the review system, and real user prompts. Can contain very personal/sensitive information https://www.404media.co/...
  • @ashleymayer Ashley Mayer on x
    This is why I always assume Sam Altman is personally reading and answering every question I ask ChatGPT (he's been great on baby stuff btw).
  • @_amirbar Amir Bar on x
    The robotics company's prisoner's dilemma: don't use GPT Astra and your competitors will show better demos and capabilities. Use it, and your data & moat will be absorbed by the next version of the LLM.
  • David Kuszmar David Kuszmar on linkedin
    The revelations of Project Lily out of OpenAI, reported by 404 Media, aren't exactly shocking as OpenAI essentially runs a giant server system hooked …
  • @evacide @evacide on bluesky
    This article makes two important points:  —  1. People using ChatGPT expose sensitive data because they think it is private  —  2. OpenAI relies hundreds of human contractors reviewing replies to improve their models  —  www.404media.co/inside-proje...
  • @josephcox Joseph Cox on bluesky
    Anthropic confirmed it is doing much the same thing too, and is using human reviewers www.404media.co/inside-proje...  [image]
  • @marypcbuk Mary Branscombe on bluesky
    I see that if we learned anything from all the times that, off the top of my head, Amazon and Apple and Facebook did this, the state of US politics prevented those lessons from actually changing anything  —  or is what we learned that a fine is a price?  [embedded post]
  • @josephcox Joseph Cox on bluesky
    This is what the workers see when they open up their dashboard: a list of “tasks”, each containing a real ChatGPT user's prompt.  They see the prompt, have to summarize what they think the person is asking for.  Includes prompt, chat memory, etc www.404media.co/inside-proje...  […
  • @josephcox Joseph Cox on bluesky
    There is a lot of new info in this article, based on documents, Slack chats, instruction guides, and me seeing real ChatGPT user prompts in the review system.  Firstly, a worker we spoke to thinks ChatGPT users have no idea their chats may be read by someone else www.404media.co/…
  • r/technology r on reddit
    Inside ‘Project Lily’: The Humans Reading Your ChatGPT Chats
  • r/BetterOffline r on reddit
    Inside ‘Project Lily’: The Humans Reading Your ChatGPT Chats
  • r/ChatGPT r on reddit
    Inside ‘Project Lily’: The Humans Reading Your ChatGPT Chats