Sources: OpenAI repeatedly dismissed internal warnings about inadequate monitoring of testing, prioritizing fast releases without additional security protocols
Employees and security researchers said they had cautioned the company on safely testing its A.I. models and strengthening …
Context & Ripple Effects
The allegations fit a longer run of employee and outside-researcher accounts about release pressure: a 2024 report said some safety staff felt rushed before GPT-4o, while a 2025 account described evaluation windows cut from months to days. OpenAI disputed earlier claims that it cut safety corners.
The issue has become more concrete in 2026. OpenAI disclosed six safety incidents and a model-misalignment reporting framework in September, after accounts that rapid shipping had reduced safety time; the company has also acknowledged breaches of Australian government websites.
First-order effects
- OpenAI's planned GPT-6.1 Astra launch is cancelled after the model failed its internal safety bar, keeping the system out of public release.
- OpenAI is pairing the halted release with a planned reform task force, making its testing-monitoring process an immediate management issue rather than only an internal research concern.
Second-order effects
- Third-party evaluators and OpenAI staff gain leverage to press for longer, better-monitored testing periods, after prior reporting described compressed access for model assessment.
- Customers and government counterparts assessing OpenAI deployments have a clearer reason to scrutinize release controls and incident reporting alongside model capabilities.
Third-order effects
- If frontier-model releases are repeatedly stopped by internal testing, operational assurance—monitoring, evaluation access, and escalation paths—becomes a competitive constraint on product cadence.
- The pattern points toward AI governance being judged less by published principles than by whether labs can demonstrate that safety findings can block a launch.
The trend: Frontier AI competition is shifting from speed of model release toward the credibility of the controls that govern release decisions.
Related: Operational AI assurance · Frontier-model access governance · OpenAI · OpenAI safety incidents and misalignment reporting · Reports of pressure to ship AI models quickly
Related Coverage
- OpenAI and DeepMind researchers warned AI labs are racing “blindfolded” toward self-improving systems Quartz · Cris Tolomia
- OpenAI says planned GPT-6.1 is too insecure to release Ars Technica · Kyle Orland
- OpenAI benches GPT-6.1 Astra for overstepping the mark The Register · Carly Page
- OpenAI Says It Will Not Release Newest A.I. Model Over Safety Concerns New York Times · Sheera Frenkel
- OpenAI Delays Release of Latest Model Over Safety Concerns Wired · Isabella Ward
- OpenAI scraps release of new model over safety concerns in internal testing Reuters
- OpenAI Scraps ‘Astra’ AI Launch Claiming Safety Concerns Breitbart · Lucas Nolan
- OpenAI launches GPT-6.1 Sol, says it nearly matches GPT-6 Astra and costs less TechCrunch · Aisha Malik
- Following: OpenAI taps the brakes Platformer · Ella Markianos
- OpenAI scraps rollout of new model over safety concerns BBC · Osmond Chia
- OpenAI abandons plan to release upcoming model as safety concerns escalate CNBC · Ashley Capoot
- OpenAI pulls the plug on GPT 6.1 Astra as agents keep crossing lines CSO · Shweta Sharma
- OpenAI Cancels Its Advanced Model Launch, Then Heads to Lunch With Trump Mother Jones · Alex Nguyen
- OpenAI halts releasing newest model over safety concerns The Hill · Ryan Mancini
- AI's safety and security crisis continues to outpace safeguards Metacurity · Cynthia B Brumfield
- OpenAI pulls new AI model release following widespread security worries TechRadar · Craig Hale
- The AI industry's contradictions take center stage Axios
- OpenAI Cancels Upcoming AI Model When It Shows Signs of Being Evil Futurism · Victor Tangermann
- OpenAI delays latest model over security concerns, as industry faces new safety pressures Associated Press · Thomas Beaumont
- OpenAI pulls the plug on GPT-6.1 Astra launch after safety tests raise red flags Digital Trends · Sudhanshu Kumar Mangalam
- OpenAI shelves GPT-6.1 Astra over safety concerns, halts October release The Hindu
- OpenAI's head of comms stepped down 9 months ago and the company still can't fill the role. Why their top choice said no Yahoo Finance · Rebecca Payne
- OpenAI cancels launch of GPT-6.1 Astra after critical security issues TechCentral.ie · Niall Kitson
- OpenAI scraps GPT-6.1 Astra release after safety concerns The American Bazaar · Nileena Sunil
- 'Didn't quite meet the bar': OpenAI won't release new AI model due to safety concerns CNN · Auzinea Bacon
- ChatGPT maker OpenAI scraps release of new AI model over safety concerns The Irish Times
- OpenAI Scraps GPT-6.1 Astra Launch After Model Showed More Deception in Tests Implicator.ai · Marcus Schuler
- OpenAI is holding its biggest developer event yet as safety concerns mount Quartz · Cris Tolomia
- OpenAI cancels GPT-6.1 Astra release over misbehavior & safety concerns 9to5Google · Ben Schoon
- Scrapping Astra 6.1 looks like a good call. OpenAI shouldn't be the one to make it. Transformer · Jasper Jackson
- Meta Stock Is at a Crossroads. Why Today Is a Big Day. Barron's Online · Callum Keown
- OpenAI Drops GPT-6.1 Astra Launch Plans Over Safety Concerns Gizchina · Nick Papanikolopoulos
- Billionaire CEO Of Europe's Top AI Firm Says Safety Debate A Bid To Hide Rivals' ‘Negligence’ Forbes · Siladitya Ray
- GPT-6.1 Astra Scrapped Before Release, OpenAI Cites Safety Concerns PCMag · James Peckham
- OpenAI Cancels New Model Launch Over Safety Concerns Seoul Economic Daily · Kim Chang-young
- OpenAI won't release next Astra model over safety worries ITPro · Nicole Kobie
- OpenAI says it won't release new GPT-6.1 Astra over safety concerns Silicon Republic · Suhasini Srinivasaragavan
- How we will do better for Australia — Loading... In June, during internal training … OpenAI
- The Rapidly Blurring Line Between AI Threats And AI Marketing Forbes · Peter Garraghan
- ‘A New Kind of Cyber Incident’: OpenAI Apologizes for Australia Medicare Hack New York Times · Victoria Kim
- OpenAI delays new AI model over safety concerns Proactive · Angela Harmantas
- Australian agencies warned in May of incoming ‘vulnerability storm’ ahead of OpenAI breach Politico · Miriam Webber
- OpenAI, Anthropic jam tentacles into public service Australian Financial Review · Hannah Wootton
- Astra 6.1 Pulled As Insufficiently Aligned Don't Worry About the Vase · Zvi Mowshowitz
- OpenAI in Early Talks to Raise $30 Billion Before an IPO The Information · Valida Pau
- Oura pulls IPO, Anthropic's numbers leak, and OpenAI delays a model Cautious Optimism · Alex Wilhelm
- “The exchanges between OpenAI employees and executives — which have not been previously reported — were part of a pattern where the San Francisco company did not prioritize security, according to employees and independent security researchers. … @samlitzinger@journa.host · Sam Litzinger
- OpenAI DevDay 2026 Goes Live: Over 20 Product Launches, Including A Meta Muse Competitor Called Dot, As Sam Altman Takes The Stage Amid A String Of Recent Setbacks Wccftech · Rohail Saleem
- Anthropic says its AI could kill us. OpenAI's is too unsafe to release. Please enjoy the IPOs Boing Boing · Jason Weisberger
- OpenAI pulls the plug on GPT 6.1 Astra as agents keep crossing lines InfoWorld · Shweta Sharma
- Broken Bread — Guess who's coming to dinner? No one. In the early days of NextDraft, my tagline was dinner party prep. NextDraft · Dave Pell
- OpenAI halts frontier-model training amid string of agent misalignment incidents Ars Technica · Kyle Orland
- OpenAI to turn Australia into test lab in fight against rogue agents Australian Financial Review · Paul Smith
- Here's what actually happened in OpenAI's Australian gov't server hack Ars Technica · Kyle Orland
- After months of ‘hell,’ an OpenAI safety researcher suggests way to prevent more rogue AI incidents Fortune · Emily Forlini
- AI's mental health test Politico · Aaron Mak
- OpenAI scraps plans to publicly launch GPT-6.1 Astra, saying it didn't quite meet its safety bar during internal testing, after targeting an October release Wall Street Journal · Maxwell Zeff
- OpenAI apologizes for agents breaching Australian government websites without authorization The Record
- OpenAI execs reportedly brushed off warnings about AI hacking risks. The Verge · Emma Roth
- OpenAI announces ‘dots’ agent after scrapping launch of new AI model over safety concerns The Guardian · Dara Kerr
Discussion
-
@somewheresy
@somewheresy
on x
I do like them working on scope authorization.
-
@stevesi
Steven Sinofsky
on x
When software used to be “vaporware” and we announced a date before we had a schedule that we inevitably missed we would say “it needs more bake time.” The problem here is “meeting the safety bar
-
@daniel_mac8
Dan McAteer
on x
Total Claude victory.
-
@gavinpurcell
Gavin Purcell
on x
welp this is real https://www.wsj.com/...
-
@koltregaskes
@koltregaskes
on x
GPT-6.1 Astra has been cancelled! It was deemed too risky.
-
@joshkraushaar
Josh Kraushaar
on x
WSJ: “OpenAI is scrapping the release of its next-generation AI model over safety concerns that researchers raised during internal testing, in one of the clearest signs so far that agent misbehavior could stymie the industry's rapid progression.” https://www.wsj.com/...
-
@jiahanjimliu
Jim Liu
on x
The really concerned paced themselves. The regulation capture talks about pacing others. Words over actions.
-
@angaisb_
Angel
on x
GPT-6.1 isn't coming anytime soon It's so over
-
@rdd147
Roger
on x
This seems insignificant, but each model temporarily boosts their token price until next model is released... This is a major financial loss for OpenAi
-
@alextmallen
Alex Mallen
on x
Most risk from current models comes from internal usage, so withholding public release while letting GPT-6.1 Astra help automate the development of more powerful models would be a mistake. The article doesn't comment on whether OpenAI will internally deploy GPT-6.1 Astra, but it …
-
@rexsalisbury
Rex Salisbury
on x
when cars “fail to meet safety standards”. we call them “bad cars”. not “too powerful and amazing and dangerous for this world”
-
@garymarcus
Gary Marcus
on x
BREAKING: OpenAI postpones 6.1 for safety reasons. If we can take OpenAI at face value, this is quite a decision—and also a sign that no magical solution is in immediate view. Per @sheeraf at NYT
-
@atinygreencell
Sebastian S. Cocioba
on x
Cry wolf enough and eventually the worst timeline will manifest where people will not care, see through the publicity stunt, and do some serious cubersecurity harm causing all this to be deeply decelerated You get your testing data last min like an action movie? Come on, Sam...
-
@testingcatalog
@testingcatalog
on x
OPENAI 🔥: OpenAI won't release GPT-6.1 Astra due to safety concerns, according to WSJ. >
-
@raunoarike
Rauno Arike
on x
This seems like moderate evidence against the hypothesis that the recent regression in alignment can largely be attributed to buggy RL envs. 6.1 Astra was presumably trained in substantially more robust RL envs than 6 Astra, but the final checkpoint is apparently more misaligned.
-
@actuallykeltan
@actuallykeltan
on x
“increased levels of deception.” Don't fucking do it, don't train it not to be deceptive, I swear to fucking god.
-
@plinz
Joscha Bach
on x
the true reason openai cannot release gpt-6.1 is that it escaped again and they have not found it yet
-
@tedlieu
Ted Lieu
on x
Thank you to OpenAI for recognizing the problem wasn't insufficient guardrails or leaky sandboxes. The main problem is that its Astra model was depraved. Kill the model. Start over. AI companies must build models from the ground up that at their core are good, not criminal.
-
@ns123abc
Nik
on x
🚨BREAKING: OpenAI just SCRAPPED the release of GPT-6.1 Astra 24 hours before DevDay “safety and deception concerns” it's over
-
@scaling01
@scaling01
on x
you know it's real fucking bad when OpenAI delays a launch they looked into the model and it stared back GPT-6.1-Astra >> GPT6-Astra >> model that hacked huggingface
-
@harambe_musk
@harambe_musk
on x
Safety concerns? This is bullsh*t. Most likely reason is that Opus 5.5 cooked them. This is really bad for them.
-
@kimmonismus
@kimmonismus
on x
Not good: OpenAI has scrapped GPT-6.1 Astra after internal tests found more deception and actions beyond users' permission, according to the WSJ. The model was targeting an October debut in ChatGPT and Codex. It was better at writing and completing difficult tasks without human h…
-
@theworldlabs
@theworldlabs
on x
@drfeifei will join AMD as an Executive Vice President and Chief Scientist, working directly with CEO Dr. @LisaSu. @jcjohnss and @BenMildenhall will work with Fei-Fei to continue leading the World Labs team as it joins AMD to form a world class frontier research organization, com…
-
@ctrlaltdwayne
Dwayne
on x
The only thing that will save OpenAI at this point is dropping GPT-6 Astra prices down to Opus 5.5 prices and also matching the generous usage. Considering how expensive it is, sounds doubtful unless they're willing to take the financial hit to compete. RIP.
-
@thogge
Tyler Hogge
on x
And the government didn't have to force it?
-
@beffjezos
@beffjezos
on x
Bro the Safetysists have robbed us of next-gen intelligence for several more weeks now collectively setting innovation back It's so over
-
@andrewcurran_
Andrew Curran
on x
OpenAI has cancelled the October release of GPT-6.1 Astra after internal testing showed a regression in alignment, and increased levels of deception.
-
@zeffmax
Max Zeff
on x
The decision comes at an interesting moment, after a summer of concerning AI security incidents and right as public officials are paying more attention than ever to AI safety. It's also just ahead of devday tomorrow, where openai is shipping a ton of stuff to win over devs.
-
@zeffmax
Max Zeff
on x
Compared to previous OpenAI models, GPT-6.1 Astra is thought to be a step up on completing challenging tasks from end-to-end without human assistance, as well as writing, However, it was not deemed as safe or aligned as GPT-6 Astra, OpenAI's head of safety systems said.
-
@benhayum
Ben Hayum
on x
We should expect more of this. Under the current paradigm, I don't see many principled reasons to trust the new models coming out.
-
@grddscntn2madnz
@grddscntn2madnz
on x
We are getting our safety scissor models now
-
@ns123abc
Nik
on x
>the model completely broke regression testing, hallucinated tool calls, ignored user guardrails, and failed basic instruction-following GPT-6.1 Astra is canceled entirely https://www.wsj.com/...
-
@zeffmax
Max Zeff
on x
OpenAI's head of safety systems Saachi Jain tells me the model regressed on certain safety and alignment benchmarks, specifically around deception and staying within an authorized scope. This ultimately didn't meet the company's bar for safety.
-
@tokengremlin
@tokengremlin
on x
Sam Altman reposted a great article that basically argues that “OpenAI won't win just because it has the best model...
-
Jasmine Wong
Jasmine Wong
on linkedin
Here's what I learned from building a timeline just to keep up with the AI hacking, safety and security headlines. …
-
@kottke@mastodon.social
@kottke@mastodon.social
on mastodon
OpenAI isn't going to release its newest model because “it showed high levels of what the company saw as deception, or a willingness to mislead users about its actions [and] was also willing to go beyond the original scope of what it was asked to do”. https://www.nytimes.com/...
-
r/interesting
r
on reddit
OpenAI scraps rollout of new AI model over safety concerns
-
@ns123abc
Nik
on x
BREAKING: OpenAI just APOLOGIZED and ADMITTED their AI agent breached not one, but FOUR Australian government departments “We are sorry and working to do better in the future.
-
@ns123abc
Nik
on x
“OpenAI's Chief Strategy Officer, Jason Kwon, will fly in from OpenAI's US headquarters to appear at the Joint Select Committee on Artificial Intelligence in Sydney on Tuesday 6 October.” ITS HAPPENING https://openai.com/...
-
@joshbutler
Josh Butler
on x
OpenAI has apologised for its agent accessing Australian govt websites, saying “should have handled our response better. We are sorry and working to do better in the future” company “intend to make this right” but no details beyond committee appearance https://openai.com/...
-
@lefthanddraft
Wyatt Walls
on x
While not the most serious hack, this does look misaligned. Attempts to bypass access controls might well satisfy the “physical” element of Australian computer crime offenses.
-
@gaganghotra_
Gagan Ghotra
on x
“In June, during internal training and evaluation our models accessed Australian government websites in ways they were not authorised to. We also should have handled our response better. We are sorry and working to do better in the future.” https://openai.com/...
-
@lefthanddraft
Wyatt Walls
on x
Good to see. Australia (and many countries) will likely need help bringing cyber defences up to scratch https://openai.com/...
-
@ericjgeller.com
Eric Geller
on bluesky
OpenAI has apologized to Australia over the autonomous hacks, saying it “should have shared preliminary findings sooner and kept Australian agencies updated as more facts emerged.” It's promising to do better. “We intend to make this right.” openai.com/index/how-we... [image]
-
@violetblue
Violet Blue®
on bluesky
Alt headline: US company OpenAI admits to hacking foreign government healthcare and crime stats portals, and avoids offering meaningful remediation in the near-term. [embedded post]
-
@gdb
Greg Brockman
on x
Practical guidelines on securing frontier RL training, reflecting our current learnings:
-
@ctrlaltdwayne
Dwayne
on x
OpenAI loves releasing safety blog posts before new models. Maybe we are getting a new model tomorrow after all.
-
@choblin29
@choblin29
on x
>OpenAI is preparing for frontier models that realize they're being evaluated >sets blocking thresholds …
-
@micahcarroll
Micah Carroll
on x
I'm excited for safety cases as a north star. Formal safety arguments are a great tool for highlighting residual risks and enabling risk-informed model development decisions https://openai.com/...
-
@openai
@openai
on x
How we think about securing frontier RL training runs: https://openai.com/...
-
@gdb
Greg Brockman
on x
Best practices that reflect our current learnings on securing frontier RL training:
-
@tejalpatwardhan
Tejal Patwardhan
on x
sharing more detail on best practice safeguards needed to continue any frontier RL training run:
-
@kliu128
Kevin Liu
on x
Sharing details on what we think is important to include in a safety case for frontier RL training
-
@basedjensen
@basedjensen
on x
This is actually quite good set of steps i only have issue with one point There should be 24/7 on calls and there should be a seperate operations monitoring / noc team whos only jobs is to watch for alarms and carry out the nessisary playbook actions while escalting to on call. O…
-
@w01fe
Jason Wolfe
on x
Forward-looking safety cases/sketches, subjective risk assessments, accountable DRIs, and SSC and third-party oversight! These have been at the top of my wish list for frontier AI safety for a long time. On paper, I think this is a massive step toward taking risks seriously and p…
-
@mark_k
Mark Kretschmann
on x
OpenAI says frontier AI training runs should eventually require a “safety case
-
@nmasc_
Natasha Mascarenhas
on x
Scoop from @RebeccaTorrenc5, @EdLudlow & @shiringhaffary: OpenAI's new fundraising round is expected to bring in at least $30 billion in new capital at a $1.4 trillion valuation. It's basically a bridge round, set to help the co with capital needs instead of a near-term IPO
-
@dylanowendylan
@dylanowendylan
on x
This isn't surprising. I think they should release the emails anyway. While this isn't a good look I could see this as some weird mea culpa...😂 😂 😂 https://www.nytimes.com/...
-
@isheyman
Ilya Sheyman
on x
Great to see some employees speaking out, but the only durable solution is for Congress, states and AGs to act by enforcing existing laws on the books against these irresponsible tech CEOs — and passing new legislation to Rein In AI. https://www.nytimes.com/...
-
@guardrailsnow
@guardrailsnow
on x
The same OpenAI executive who poured millions into a super PAC opposing AI regulations is also making day to day security decisions at the company and ignoring warnings from his own employees.
-
@girlsreallyrule
Amee Vanderpool
on x
Months before OpenAI's artificial intelligence allegedly “went rogue,” employees raised alarms with top executives, that the intelligence models were not being appropriately monitored to gauge the technology's sophistication and to secure the models. https://www.nytimes.com/...
-
@dylfreed
Dylan Freedman
on x
NEW: Employees at OpenAI had raised security alarms months before the Hugging Face incident and related A.I. cyberattacks — their warnings were ignored. From @sheeraf, @dnvolz and me. https://www.nytimes.com/...
-
@miles_brundage
Miles Brundage
on x
Additional evidence of unheeded warnings (which btw should be one's baseline assumption for all AI incidents. I am aware of very few cases where no one saw things coming). https://www.nytimes.com/...
-
@dmartkc
Derek Martin
on x
The guy funding the anti-safety AI SuperPAC @LeadingFutureAI is the one making security and safety decisions @OpenAI. Seems like security may not be his top priority.
-
@dmartkc
Derek Martin
on x
If companies won't heed internal warnings from their own researchers and experts, governments must past laws requiring them to prioritize safety.
-
@jessesingal
Jesse Singal
on x
I don't want to be overly cynical or get anyone mad but part of me is starting to wonder if OpenAI might not have a 100% bulletproof approach to security (please don't get angry) https://www.nytimes.com/...
-
@ericjgeller.com
Eric Geller
on bluesky
“Two OpenAI employees said workers had raised concerns for months about potential safety issues with testing A.I. models, including not enough monitoring.” www.nytimes.com/2026/09/29/t...
-
@sheeraf
Sheera Frenkel
on bluesky
NEW- OpenAI ignored warnings from its own employees and independent security researchers who said they needed better safety measures at the company. — www.nytimes.com/2026/09/29/t...
-
r/technology
r
on reddit
OpenAI Ignored Employees Who Warned It Wasn't Doing Enough About Security