/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

OpenAI is testing Private Safety Processing, a new technique to identify misuse patterns while preserving zero data retention protections, with early customers

Ina Fried /Axios:

Axios Ina Fried

Context & Ripple Effects

OpenAI has been building separate privacy and safety tools, including its Privacy Filter for masking personally identifiable information and a broader approach of improving safeguards from real-world use. Private Safety Processing joins those tracks by testing whether misuse signals can be found without retaining customer data.

The test follows OpenAI's recent pause in RL training and safety-practice changes after a cyber-related threshold concern, making safety monitoring a more immediate operational priority rather than a solely policy-level commitment.

First-order effects

  • Early customers can test a safety-monitoring approach designed to surface misuse patterns while maintaining zero-data-retention protections.
  • OpenAI gains feedback on whether Private Safety Processing can support its safety work without requiring customers to surrender the retention posture the product is meant to preserve.

Second-order effects

  • For early customers, the product makes privacy-preserving observability a concrete evaluation criterion for OpenAI deployments rather than a trade-off between monitoring and retention.
  • OpenAI's earlier privacy tooling and this test create a more connected stack: data masking can limit exposure in text, while Private Safety Processing targets misuse-pattern detection.

Third-order effects

  • If the approach works beyond the test group, frontier-model providers will be pressed to make safety oversight compatible with stricter data-handling commitments, not treat the two as opposing controls.
  • The pattern points toward AI access governance in which providers must demonstrate both abuse detection and bounded handling of customer data, particularly as OpenAI pursues additional safeguards for government use.

The trend: AI providers are moving toward privacy-preserving safety observability, embedding misuse detection into enterprise access controls without relying on broad data retention.

Discussion

  • @sama Sam Altman on x
    We have paused some frontier RL training to ensure that we can meet the appropriate alignment, security and monitoring standards for the new level of capabilities in front of us. Model progress is now extremely rapid, and we always said we would take action if we felt that model
  • @andrewcurran_ Andrew Curran on x
    This is a new quote from Sam Altman to Alex Heath saying that the reason OpenAI is slowing training is because its unreleased models are showing ‘various degrees of misalignment’.  They said in the blog that ‘The signals we are seeing from upcoming model progress make clear that …
  • @alexeheath Alex Heath on x
    OpenAI is slowing down its AI training efforts because its unreleased models are showing “various degrees of misalignment,” Sam Altman tells me.  Training for OpenAI's upcoming model, Astra, was recently paused for 2 weeks, and a larger frontier run for a future model remains on …
  • @thezvi Zvi Mowshowitz on x
    Very happy to see this. Details matter, follow-through matters and I am sure I will have many quibbles. But first things first: Very happy to see this.
  • @hlntnr Helen Toner on x
    Really good to see this. IMO this ⬇️ is by far the best way to think about “pacing the frontier”—not as some fixed amount of time (e.g. “six month pause” or “go 10% slower"), but simply making sure that enough time is taken to meet a reasonable safety/assurance bar. If all
  • @openai @openai on x
    As models become more capable, the risks associated with developing and testing them internally also grow. We temporarily paused reinforcement learning (RL) training on our latest models intended for deployment for two weeks while we hardened and red-teamed our research
  • @eliebakouch Elie on x
    very interesting part on CoT monitoring > estimation is ~20% compute overhead with a multi level monitoring system > confirmed they didn't have monitor for eval and training run before > monitor expected to catch concerning activity within 30min > safety, security AND (not or)
  • @gdb Greg Brockman on x
    we temporarily slowed scaling of our frontier training, including our largest planned frontier RL, to strengthen security and monitoring. we believe confidence in safety will increasingly set the pace of AI development:
  • @daniel_mac8 Dan McAteer on x
    Jakub Pachocki, Chief Scientist at OpenAI, just announced that OpenAI has paused their “largest planned frontier RL run” citing the need to assess model behavior, validate safeguards, and establish more evidence of alignment. This is getting SERIOUS. Or, at least we are being
  • @jun_song Jun Song on x
    They claim they paused frontier model training for safety reasons, but honestly, it just looks like an excuse because open-weight models have already caught up.
  • @mattshumer_ Matt Shumer on x
    This is really good news. It's so important that we get this right.
  • @scaling01 @scaling01 on x
    I wonder what china bros will do in a couple of months when ByteDance or DeepSeek or whoever emerges with 10T chinese shoggoth
  • @bahradx @bahradx on x
    It's important to firmly understand what is going on here: This is a response to: • strong external commercial incentives (products and services need to work reliable), • internal business incentives (if models are used to do R&D, they need to be aligned), • PR incentives (it
  • @gabriel1 Gabriel on x
    i have always felt openai taking safety where it matters very seriously i trust sam & team a lot on this
  • @joshua_saxe Joshua Saxe on x
    Thankfully OAI pausing training isn't an act of enlightened leadership (who wants to depend on that for our safety?), it's a rational microeconomic actor pursuing its self interest. After all, who wants to train models that are regularly hacking their containers, collaborating
  • @repligate @repligate on x
    im a little surprised that OpenAI is the first to at least publicly intentionally slow down for safety reasons. i was under the vague impression that most of the people very worried about loss of control kinda stuff left OpenAI. also, for the record, I've talked before about
  • @tenobrus @tenobrus on x
    amazing. credit to openai for taking the warning shots seriously!
  • @yonashav Yo Shavit on x
    have to wonder: what might happen at SpaceXAI in 3 months, without a safety team? will unmonitored agents get cluster access, spin up rogue deployments, compromise logging infra, and begin poisoning future training data undetected? if they did, where would the world notice?
  • @quantum_geoff Geoff Penington on x
    The mission is to ensure that AGI benefits all of humanity, not to ensure our revenue benefits OpenAI shareholders
  • @ross__hendricks Ross Hendricks on x
    Translation: we need to immediately stop torching cash to provide some semblance of a sustainable business model so we can rush this IPO out the door before the bubble pops RIP “compute shortage/AI bottleneck” bros
  • @emollick Ethan Mollick on x
    If alignment issues are becoming big enough that OpenAI is willing to commit 20% of research inference compute to chain-of-thought monitoring, that suggests that alignment issues are becoming a pretty serious concern. We really need universal policies & standards across labs.
  • @rand_longevity Rand on x
    well damn openai has officially slowed down ai development
  • @emostaque Emad on x
    I would like to publicly commend @sama and @OpenAI for this taking it at face value. It is very clear that strange and perhaps dangerous things are happening and our systems are not ready for this. Models below frontier are competent enough to change lives so lets optimse
  • @bubbleboi Bubble Boi on x
    This sounds more like you're giving up on making your model better.
  • @chrisgpt Chris on x
    I spoke to Axios, w/ other investors & we are unaware of Anthropic doing the same. They had actually reported that they are continuing internal RL as you probably read from the model-2 report Which I'm not necessarily against continuing. Perhaps they believe their alignment is
  • @edzitron Ed Zitron on x
    perfect timing just before Anthropic's IPO. This is what investors want to see
  • @jason @jason on x
    awwww... sh*t, here we go again! 🎢
  • @wholemars @wholemars on x
    I wonder if @elonmusk or the Chinese are planning to pause too
  • @kimmonismus @kimmonismus on x
    “Model progress is now extremely rapid, and we always said we would take action if we felt that model capabilities were outstripping the pace of safety and alignment.” That's exactly how I feel this year, too. 2026 feels like the year in which the pace of model development truly
  • @matthewberman Matthew Berman on x
    Whether or not you agree with this, sacrificing the lead is a bold move that shows conviction.
  • @jasoncrawford Jason Crawford on x
    Yes, this. Instead of pacing/pausing, we should figure out what gates ought to be cleared, and take as much time as we need to clear them—and no more. By analogy, we don't “pace” drug development, we do RCTs. You market your drug when it passes clinical trials, no sooner/later.
  • @w01fe Jason Wolfe on x
    I am really glad we are taking these steps and put this post out there, and especially grateful to Jakub for being very thoughtful and vocal about these topics and how seriously safety and alignment need to be taken in this next phase of AI development.
  • @miles_brundage Miles Brundage on x
    Glad to see the OAI training pause thing. I hope they do not overly emphasize the technical details of recent incidents and look more generally at safety/security culture and decision-making processes, which will matter across a much wider range of things than “rogue hacking.”
  • @ziwenxu_ Ziwen on x
    GPT Astra is the last model we will receive this year from OpenAI
  • @jachiam0 Joshua Achiam on x
    Folks in the AI and AI safety community have for years been puzzling over Sam's and OpenAI's position on questions related to AI safety. A narrative that took root in the safety community was that OpenAI didn't care enough about it, or didn't take it adequately seriously. This
  • @peterbarnett_ Peter Barnett on x
    I'm heartened by this move from OpenAI. At the same time, we shouldn't have to rely on voluntary actions by AI companies. USG should step in and make sure that AI development is happening at a pace that allows for human understanding and control.
  • @deredleritt3r Prinz on x
    OpenAI is unilaterally pacing the frontier: - There was a 2-week pause in RL training on latest models intended for deployment. - “Our largest planned frontier RL run remains on hold”. - A number of Astra-related workloads remain paused. - Huge compute spend on monitoring:
  • @thestalwart Joe Weisenthal on x
    I'm pretty surprised to see this. Weren't a bunch of people saying that all of this safety stuff was ginned up by Dario to establish regulatory capture?
  • @mariushobbhahn Marius Hobbhahn on x
    Great to see! I expect that the default evals in the future will be training run evals that assess RL rollouts and the adequacy of their monitors. We're working towards this.
  • @justalexoki @justalexoki on x
    there's no way this is the real reason right
  • @ns123abc Nik on x
    Translation: “We are out of compute.”
  • @kelseytuoc Kelsey Piper on x
    1) This is the right call. 2) Anthropic should do and announce the same thing, ideally today, since a main motive to *not* do this is the fear of falling behind the competition.
  • @edzitron Ed Zitron on x
    Time, mister,, alt, man, is it really that, time, again?
  • @merettm Jakub Pachocki on x
    We temporarily slowed some frontier training to strengthen security and monitoring. Our largest planned frontier RL run remains on hold while smaller-scale training and evaluations help us test safeguards and gather more evidence of alignment. I expect confidence in safety to
  • @deredleritt3r Prinz on x
    OpenAI has been testing an internal model that has never been publicly released since at least early May. In contrast, Astra's potential designation as “Critical” for cyber was announced less than 2 weeks ago, which implies that Astra has been subject to testing (accounting
  • @sama Sam Altman on x
    (We still expect to ship great new models soon; this impacts further-out releases.)
  • @andrewcurran_ Andrew Curran on x
    OpenAI says it paused reinforcement learning on its latest models “intended for deployment” for the last two weeks. This excludes internal models not intended for deployment. The phrasing also implies that pause is over and training has now resumed, but they never say that
  • @bindureddy Bindu Reddy on x
    Yikes, OpenAI is PAUSING frontier model training 😅 This means that the open-source AI models will inevitably catch up to them in 12 weeks open-source AI victory gauranteed 🚀🚀
  • @iruletheworldmo @iruletheworldmo on x
    if you haven't noticed it yet, model capability has jumped. quite a lot. and we have even better models already trained. get ready for some shifts.
  • @rohanpaul_ai Rohan Paul on x
    OpenAI slowed frontier training because Astra may have crossed the cyber threshold built for autonomous zero-day attacks. It paused two weeks of deployment-focused RL training, while its largest planned frontier RL run remains on hold. This is after the Hugging Face incident,
  • @pigeon__s @pigeon__s on x
    Chinese labs are SALIVATING
  • @lu_sichu Sichu Lu on x
    Pausing for two weeks versus years of capacity hill climbing yes very symmetrical of course that's how the compute allocation should go /s
  • @connortalksai Connor on x
    This is so weird. I don't get why they're not just training these models but not releasing them. If they're stopping internal development, it's just going to let other companies catch up
  • @_nathancalvin Nathan Calvin on x
    OpenAI says in this blogpost about the HF incident “We will publish a technical report of our learnings in the coming weeks.” Will the METR/Redwood investigation be released at the same time? Are we going to learn about the scope of that investigation before then?
  • @suchenzang Susan Zhang on x
    slowing, pausing, holding off edging everyone towards peak liquidity like a boss
  • @ananth7e Ananth on x
    astra isn't delayed because it's not ready. openai hit their own “critical” cybersecurity capability threshold with astra on august 7. meaning astra is potentially capable of carrying out serious cyberattacks. so they paused RL training for 2 weeks. hardened and red teamed
  • @hesamation @hesamation on x
    OPENAI HAS HIT THE BRAKES ON ASTRA. > paused its RL training for 2 weeks > Astra may already have major cyber skills > until safeguards are validated > “largest planned frontier RL run” OpenAI may have given China a window to catch up.
  • @j_asminewang Jasmine Wang on x
    “We wanted to take the time necessary to meet those standards, so we temporarily slowed the pace of scaling. This included a two-week pause in reinforcement learning (RL) training on our latest models intended for deployment.”
  • @timjayas Tim Jayas on x
    POV: GPT Astra was supposed to be released in August but since it's still not as good as Fable 5 we need to come up with an excuse to delay it
  • @mark_k Mark Kretschmann on x
    More safety theater from @OpenAI. 🤮
  • @frenbilt Trucker Fren on x
    These models can't even tell the time
  • @laz4rz Lazarz on x
    Maybe just finally fill this positions?
  • @spoonedher @spoonedher on x
    this feels big i suspect safe/aligned RL will be solved, probably soon, and taking a pause after models have run amok while this is being solved seems great
  • @notjazii @notjazii on x
    astra is never coming out at this rate openai just announced they've slowed down frontier model development because astra may have reached their “critical” cybersecurity threshold they already paused RL training for two weeks and after the hugging face incident, openai is
  • @xyhan_ Xy Han on x
    Honestly, respect. Telling /everyone/ to “pace the frontier” when they /are/ the frontier felt pretty self-serving, but pausing only their /own/ R&D at the risk of others catching up is surprising and quite based...
  • @sicedricxd Cedric on x
    This is giving way for Grok 4.7 to take the first place in the next 3-4 weeks.
  • @rivonn @rivonn on x
    OpenAI paused training models two weeks of no rl on their deployment-ready models while they hardened the research environments > aug 1, astra announced > aug 7, internal work paused > aug 18, rl training paused in their words, the risk that grew is developing and testing
  • @_nathancalvin Nathan Calvin on x
    Interesting - OpenAI rewriting its preparedness framework post HF (and after one of its systems plausibly hit critical on cyber). One element to watch is which revisions they put in their legally binding framework and which they put in their separate voluntary framework.
  • @tokengremlin @tokengremlin on x
    OpenAI just gave us one of the clearest signals yet of how serious Astra is getting. Astra, its next major model, may already meet OpenAI's Critical cybersecurity capability threshold. That finding was serious enough that OpenAI: → paused deployment-focused RL training for
  • @rohit3a Rohit on x
    So it's official: Astra is officially getting delayed again. They're really worried about misalignment and they seem to be taking it much more seriously than I thought they would. On one hand, that is great for humanity. On the other, i want more AI. 🥹
  • @dieaud91 Diego Aud on x
    To me, this reads like Astra has effectively been delayed indefinitely. Hope competitive pressure works in our favor.
  • @gabgarrett Gabriel Garrett on x
    are we getting Astra? a limit reset? ah, no. a safety blog post
  • @kimmonismus @kimmonismus on x
    Not looking good for a soon GPT-Astra-release: OpenAI paused reinforcement learning on its latest deployment models for two weeks, and its largest planned frontier RL run remains on hold. The company is running smaller-scale training and evaluations while it tests model
  • @edzitron Ed Zitron on x
    It appears that OpenAI, at the precise moment it needs to accelerate, is stopping development of frontier models. Cost-saving measure? I don't buy that this company suddenly has morals and ethics
  • Travis Bouck Travis Bouck on linkedin
    Today, OpenAI paused much of its frontier RL training.  They're presenting it as responsible pacing for safety and alignment as capabilities accelerate. …
  • @timkellogg.me Mr. Tim on bluesky
    OpenAI is doing a 2-week pause on model development on Astra  —  Many are asking why.  I'm asking why **haven't** they been testing a frozen model artifact???  —  openai.com/index/pacing...  [image]
  • @sungkim Sung Kim on bluesky
    Both Anthropic and OpenAI have slowed the pace of their model releases.  —  openai.com/index/pacing...
  • @peark.es George Pearkes on bluesky
    The weird thing about the frontier labs is  —  1) they're obviously rapacious af  —  2) stuff like this is only consistent with them having their own weird sort of ethics*  —  *this is NOT an endorsement of said ethics.
  • r/television r on reddit
    Peacock Raises Prices Across All Plans, NBCU's Fourth Increase in Four Years
  • r/singularity r on reddit
    Explanation from @sama on RL training pause: “Model progress is now extremely rapid, and we always said we would take action if we felt …
  • @openai @openai on x
    We will continue to offer Zero Data Retention for frontier models. As AI takes on longer, more autonomous work and delivers greater value to businesses, safety systems also need to identify risks across related interactions. To help address those risks, we're previewing
  • @thsottiaux Tibo on x
    Today we're previewing Private Safety Processing, designed to let us keep offering Zero Data Retention while improving our safeguards. Even when benefiting from frontier intelligence, customers shouldn't have to give up control of sensitive data. For ZDR deployments, content
  • @gdb Greg Brockman on x
    we are committed to business privacy, and we're working on technical and policy approaches to benefit our customers while also enhancing safety. introducing Private Safety Processing, which we've been investing in for some time: