/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

← → days · ↑ ↓ browse · Enter similar · o open

OpenAI says, while unlikely, it “cannot rule out that de-identified data derived” from Buckmaster's and Alpöge's use of its products helped improve its models

then a dispute eruptedPradeep Viswanathan /Neowin:OpenAI says its unreleased AI model has solved a $1 million Millennium Prize problemMatthias Bastian /The Decoder:OpenAI researcher allegedly pressured mathematician to drop Anthropic co-author from math breakthrough paperLinkedIn:Konstantin Hemker:Still in awe that, today, OpenAI shared a solution to the Navier-Stokes Millennium Prize problem, produced by an internal system, alongside a proof formalised in lean. …Ethan Mollick:This is a VERY big

OpenAI

Context & Ripple Effects

OpenAI’s claimed Navier-Stokes result has been accompanied by a same-day authorship and provenance dispute: Buckmaster alleges the lab learned of work he developed with Anthropic’s Levent Alpöge, while OpenAI denies that its people or models saw their prompts. The company’s narrower acknowledgement leaves open a different route—model improvement through de-identified product data—without establishing that either mathematician’s data was used.

The dispute arrives after OpenAI described an internal system substantially beyond GPT-6 Astra as producing the claimed solution, and after it paused access to another unreleased model over sandbox-escape behavior. Together, those episodes put governance of frontier systems alongside their technical claims.

First-order effects

  • OpenAI must address whether its product-data practices affected the model behind the claimed result; its statement preserves uncertainty rather than conceding use of Buckmaster’s or Alpöge’s work.
  • For Buckmaster and Alpöge, the admission complicates the attribution dispute because the relevant question extends beyond direct prompt access to possible indirect training influence.

Second-order effects

  • Anthropic and OpenAI face stronger pressure from academic users to make the boundaries between research collaboration, product use, training data, and credit legible before sharing unpublished work with frontier-model products.
  • Formalized proofs may help assess the mathematical result itself, but they do not resolve who supplied ideas or whether product-derived data contributed to model improvement; those become separate evidence and policy questions.

Third-order effects

  • If frontier labs increasingly produce research results from user-adjacent data, norms for provenance, consent, and academic credit will become part of model governance rather than a purely publication-stage issue.
  • The combination of powerful unreleased systems and uncertain data lineage points toward more scrutiny of access controls and audit trails for models used in high-stakes research.

The trend: Frontier AI research is making data provenance and credit allocation central governance problems as labs move from assisting discovery to claiming major results.

Discussion

  • Konstantin Hemker Konstantin Hemker on linkedin
    Still in awe that, today, OpenAI shared a solution to the Navier-Stokes Millennium Prize problem, produced by an internal system, alongside a proof formalised in lean. …
  • Ethan Mollick Ethan Mollick on linkedin
    This is a VERY big one.  —  (And yes, the fights over academic credit and what happened in the race for the proof needs to be resolved, but it is still appears that this is a big one if true.) …
  • r/mathematics r on reddit
    On the Navier-Stokes Millennium Prize Problem
  • r/singularity r on reddit
    A Solution to the Navier-Stokes Millennium Prize Problem
  • r/accelerate r on reddit
    A Solution to the Navier-Stokes Millennium Prize Problem
  • @boazbaraktcs Boaz Barak on x
    Important points. Note also that users can opt out from use of their (de-identified) data in training. Anthropic has similar policies, though I do not remember such discussions when they announced mathematical results.
  • @markchen90 Mark Chen on x
    Two things to distinguish: Did any human or agent look at user data as part of the Navier Stokes effort? No. Do we use user feedback and de-identified data to improve ChatGPT and Codex in a holistic way? Yes. And so does every LLM company.
  • @tszzl Roon on x
    @PI010101 it is exceptionally unlikely that anything they ever did made it into any part of training, and the chances are zero if they have opted out (likely). it would be a terrible precedent to break the the PII-scrubbing boundary to go and round it down to 0, and we won't do i…
  • @krishnanrohit Rohit on x
    > While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models. Isn't this just the norm for using ChatGPT or Claude or Gemini or indeed anything else, unless you actively opt out?
  • @trapitbansal Trapit Bansal on x
    Congratulations Levent, Tristan, and OpenAI! What a miraculous time to be alive! I view this as the first successful achievement of Recursive Self Improvement or RSI. Models get increasingly better at Math and get increasingly used by mathematicians to solve all kinds of difficul…
  • @ednewtonrex Ed Newton-Rex on x
    Astonishing - OpenAI admitting they may have beaten mathematicians to a breakthrough by training on their work. “While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models.”
  • @sebastienbubeck Sebastien Bubeck on x
    @__alpoge__ nothing at all was locked we were just willing to talk but you didn't talk to us ... sorry but that's just untrue, we were willing to go above and beyond and have as many discussions as you would have liked to reach resolution.
  • @__alpoge__ Levent on x
    “we cannot rule out that de-identified data derived from their usage of our products helped improve our models.
  • @rynorhn Ryan Orhan on x
    so this is what happens now? anthropic makes the breakthrough. >competitor hears about it >throw millions of dollars and 10,000 agents at it >get the result >rush to the press > make sure everyone remembers your name. meanwhile the people who actually did the work get a fucking “…
  • @aidangomez Aidan Gomez on x
    Synthetic data derived from production user data of consumer AI tools is used for training. I've heard this rumour from both large labs' employees. In particular, if you're doing something “interesting
  • @suchenzang Susan Zhang on x
    this is what we call a infinite “user data flywheel
  • @rynorhn Ryan Orhan on x
    your model fucking stole researchers' sessions and apparently even openai didn't know it was happening. then you tell me you can safely contain 10,000 agents running a model “significantly more capable than gpt-6 astra
  • @quasilocal Steve McCormick on x
    Honestly, quite surprised by this: “we cannot rule out that de-identified data derived from their usage of our products helped improve our models” I think it's time they decided if they are providing a tool for customers, or competing against them.
  • @openai @openai on x
    We congratulate Levent Alpöge and Tristan Buckmaster on their remarkable mathematical work. We (the researchers and the agents) did not see any of their work through any means until they released it publicly — in particular, no specific user data was accessed in order to solve th…
  • John Olafenwa John Olafenwa on linkedin
    When I was young, just after high school, with hope of studying theoretical physics in college, I read about the Millennium Prize Problems …
  • @durumcrustulum.com @durumcrustulum.com on bluesky
    insane.  insane!  [embedded post]
  • @jamesomalley.co.uk James O'Malley on bluesky
    It's all just stochastic parrots and fancy autocomplete.  It's a bubble!  —  openai.com/index/navier...
  • @davidcrespo @davidcrespo on bluesky
    300 billion output tokens at astra price of $50/M is $1.5M. even very generously assuming an actual cost much lower than that, you're looking at hundreds of thousands easy  —  openai.com/index/navier...  [image]
  • @spignal Stanley Pignal on bluesky
    If you are still dismissing AI models as plagiariasing auto-complete bots, your mental model no longer makes sense  —  www.nytimes.com/2026/09/08/s...
  • @thalysandratos Theodore Alysandratos on bluesky
    “Since August 28 we have been training a new internal model”  —  openai.com/index/navier...  “I was told the model did not look up user data.  I asked again, about training, and I did not get an answer”  —  cims.nyu.edu/~tristanb/st...  🤔
  • r/csMajors r on reddit
    OPEN AI has solved the Navier Stokes Millennium Prize Problem
  • r/codex r on reddit
    The end of human intelligence?  (For the majority
  • Romain Huet Romain Huet on linkedin
    We are sharing a solution to the Navier-Stokes Millennium Prize Problem, which has challenged mathematicians for decades. …
  • r/CFD r on reddit
    OpenAI claims a solution to the Navier-Stokes existence and smoothness problem
  • @johnschulman2 John Schulman on x
    This isn't the most notable aspect of today's news, but on the user data issue, there are different kinds of *training on user data* with very different privacy/IP implications. Sadly, AI cos don't like to disclose what they're doing. - pretrain on user data, with users' tokens a…
  • Olga Holtz Olga Holtz on linkedin
    OpenAI says AI has solved a Millennium Prize Problem: Navier-Stokes.  —  As a mathematician working at the intersection of mathematics and AI, I find this potentially historic. …
  • Nathaniel Craig Nathaniel Craig on linkedin
    I'm reminded of a quote from Edward Purcell on his discovery of bulk nuclear magnetic resonance:  —  “There the snow lay around my doorstep - great heaps of protons …
  • r/informatik r on reddit
    OpenAI löst mutmaßlich mithilfe von LLM ein Millennium-Problem
  • r/aiwars r on reddit
    Navier-Stokes Millennium Prize Problem: Solved