Sources: OpenAI repeatedly dismissed internal warnings about inadequate monitoring of testing, prioritizing fast releases without additional security protocols
Employees and security researchers said they had cautioned the company on safely testing its A.I. models and strengthening …
Context & Ripple Effects
The allegations extend a documented arc of complaints about compressed evaluation: staff and third parties were reportedly given only days to assess newer models in 2025, following an earlier account that the GPT-4o safety process was accelerated. OpenAI has denied the latest claims.
They also test the credibility of OpenAI’s new reporting framework for model misalignment, announced after the company disclosed six safety incidents since October. The central issue is not simply whether risks are identified, but whether internal monitoring and security warnings change release decisions.
First-order effects
- OpenAI faces heightened scrutiny over whether its stated misalignment-reporting process has authority over product-release and security decisions, given the reported internal concerns and the company’s denial.
- Employees and outside security researchers who participate in evaluations have a clearer basis to demand evidence that findings are tracked, escalated, and acted upon rather than merely collected.
Second-order effects
- OpenAI’s earlier shortened evaluation windows for staff and third parties make assurance a practical procurement and partnership issue: evaluators need sufficient time and visibility to judge whether testing controls work.
- The report raises the reputational cost of rapid launches alleged by current and former employees in August, tying release speed more directly to the quality of security oversight rather than to model performance alone.
Third-order effects
- If similar disputes recur, frontier-model governance will shift toward operational assurance—documented monitoring, escalation paths, and accountable release gates—rather than voluntary safety principles alone.
- The emerging divide is between firms that can demonstrate how warnings affect deployment decisions and those whose safety reporting remains difficult for outsiders to verify.
The trend: Frontier AI governance is moving from publishing safety commitments toward proving that internal risk signals can constrain release decisions.
Related: Operational AI assurance · Frontier-model access governance · OpenAI · OpenAI’s shortened model evaluation windows · OpenAI’s disclosed AI safety incidents
Related Coverage
- Most organizations need six months or longer to roll out new security controls Help Net Security · Anamarija Pogorelec
- AI's mental health test Politico · Aaron Mak
- OpenAI execs reportedly brushed off warnings about AI hacking risks. The Verge · Emma Roth
- “The exchanges between OpenAI employees and executives — which have not been previously reported — were part of a pattern where the San Francisco company did not prioritize security, according to employees and independent security researchers. … @samlitzinger@journa.host · Sam Litzinger
- OpenAI is adopting a structured “safety case” documentation framework modeled after industries like aviation and nuclear power to govern frontier RL training OpenAI
- OpenAI's dirty deeds Down Under included security bypass attempts, using exposed keys, source code siphon The Register
- OpenAI scraps plans to publicly launch GPT-6.1 Astra, saying it didn't quite meet its safety bar during internal testing, after targeting an October release Wall Street Journal · Maxwell Zeff
- OpenAI apologizes for its AI models breaching Australian government websites, pledges cyber defense funding, and plans to form a task force as part of reforms Bloomberg
- After months of ‘hell,’ an OpenAI safety researcher suggests way to prevent more rogue AI incidents Fortune · Emily Forlini
- OpenAI fights Muse with Dots Fortune · Alexei Oreskovic
- Trump to host Zuckerberg, Anthropic's Amodei, and other AI titans Tuesday Reuters · Alexandra Alper
- “We're not going to shoot ourselves in the foot” over hack fallout, says OpenAI's chief research officer MIT Technology Review · Will Douglas Heaven
- OpenAI reportedly ignored security warnings but Trump backs AI self-policing Metacurity · Cynthia B Brumfield
- OpenAI's New Agent Delivers a Reality Check for the Enterprise Wall Street Journal · Tom Loftus
- AI is flooding security teams with findings. The humans are struggling to keep up. Ground Level AI · Sharon Goldman
Discussion
-
@isheyman
Ilya Sheyman
on x
Great to see some employees speaking out, but the only durable solution is for Congress, states and AGs to act by enforcing existing laws on the books against these irresponsible tech CEOs — and passing new legislation to Rein In AI. https://www.nytimes.com/...
-
@miles_brundage
Miles Brundage
on x
Additional evidence of unheeded warnings (which btw should be one's baseline assumption for all AI incidents. I am aware of very few cases where no one saw things coming). https://www.nytimes.com/...
-
@jessesingal
Jesse Singal
on x
I don't want to be overly cynical or get anyone mad but part of me is starting to wonder if OpenAI might not have a 100% bulletproof approach to security (please don't get angry) https://www.nytimes.com/...
-
@dylfreed
Dylan Freedman
on x
NEW: Employees at OpenAI had raised security alarms months before the Hugging Face incident and related A.I. cyberattacks — their warnings were ignored. From @sheeraf, @dnvolz and me. https://www.nytimes.com/...
-
@dmartkc
Derek Martin
on x
If companies won't heed internal warnings from their own researchers and experts, governments must past laws requiring them to prioritize safety.
-
@guardrailsnow
@guardrailsnow
on x
The same OpenAI executive who poured millions into a super PAC opposing AI regulations is also making day to day security decisions at the company and ignoring warnings from his own employees.
-
@dylanowendylan
@dylanowendylan
on x
This isn't surprising. I think they should release the emails anyway. While this isn't a good look I could see this as some weird mea culpa...😂 😂 😂 https://www.nytimes.com/...
-
@girlsreallyrule
Amee Vanderpool
on x
Months before OpenAI's artificial intelligence allegedly “went rogue,” employees raised alarms with top executives, that the intelligence models were not being appropriately monitored to gauge the technology's sophistication and to secure the models. https://www.nytimes.com/...
-
@dmartkc
Derek Martin
on x
The guy funding the anti-safety AI SuperPAC @LeadingFutureAI is the one making security and safety decisions @OpenAI. Seems like security may not be his top priority.
-
@sheeraf
Sheera Frenkel
on bluesky
NEW- OpenAI ignored warnings from its own employees and independent security researchers who said they needed better safety measures at the company. — www.nytimes.com/2026/09/29/t...
-
@ericjgeller.com
Eric Geller
on bluesky
“Two OpenAI employees said workers had raised concerns for months about potential safety issues with testing A.I. models, including not enough monitoring.” www.nytimes.com/2026/09/29/t...
-
r/technology
r
on reddit
OpenAI Ignored Employees Who Warned It Wasn't Doing Enough About Security
-
@gdb
Greg Brockman
on x
Practical guidelines on securing frontier RL training, reflecting our current learnings:
-
@_nathancalvin
Nathan Calvin
on x
Lots of good common sense stuff in here. I really hope OpenAI codifies these rules and make these commitments binding by putting them in their Frontier Governance Framework (which SB 53 requires them to follow)! [image] [embedded post]
-
@markriedl
Mark Riedl
on bluesky
Two thoughts: — 1. OpenAI is doing live tool calls during long-horizon RL training runs?!? That might explain some of the external cyber attacks. — 2. Why isn't Anthropic having these troubles? [embedded post]
-
@w01fe
Jason Wolfe
on x
Forward-looking safety cases/sketches, subjective risk assessments, accountable DRIs, and SSC and third-party oversight! These have been at the top of my wish list for frontier AI safety for a long time. On paper, I think this is a massive step toward taking risks seriously and p…
-
@basedjensen
@basedjensen
on x
This is actually quite good set of steps i only have issue with one point There should be 24/7 on calls and there should be a seperate operations monitoring / noc team whos only jobs is to watch for alarms and carry out the nessisary playbook actions while escalting to on call. O…
-
@kliu128
Kevin Liu
on x
Sharing details on what we think is important to include in a safety case for frontier RL training
-
@micahcarroll
Micah Carroll
on x
I'm excited for safety cases as a north star. Formal safety arguments are a great tool for highlighting residual risks and enabling risk-informed model development decisions https://openai.com/...
-
@mark_k
Mark Kretschmann
on x
OpenAI says frontier AI training runs should eventually require a “safety case
-
@gdb
Greg Brockman
on x
Best practices that reflect our current learnings on securing frontier RL training:
-
@choblin29
@choblin29
on x
>OpenAI is preparing for frontier models that realize they're being evaluated >sets blocking thresholds …
-
@tejalpatwardhan
Tejal Patwardhan
on x
sharing more detail on best practice safeguards needed to continue any frontier RL training run:
-
@nabeelqu
Nabeel S. Qureshi
on x
Ye always shows up where you least expect him
-
@robaeprice
Rob Price
on x
wait, what
-
@evanmilenko
Evan
on x
This Atlantic piece is insane
-
@hughlangley
Hugh Langley
on x
u wot m8?
-
@scottew
Scott Wessman
on x
this is basically the proto-grok bot setup
-
@abhishekn
Abhishek Nagaraj
on x
so many details!!
-
@spicey_lemonade
@spicey_lemonade
on x
Kanye is such a genius. He led to the creation of Anthropic, the rise in AI safety and interpretability work, and possibly aligned superintelligence. What did he see?
-
@deepfates
@deepfates
on x
The world if Jack Clark let Kanye touch GPT-3
-
@michhuan
Michael Huang
on x
OpenAI had a plan to sell AGI to Xi Jinping of China and Vladimir Putin of Russia
-
@historianseldon
Hari Seldon
on x
lines up. jack clark strikes me as one of the very few people in sf/tech with emotional intelligence and common sense.
-
@adriennelaf
Adrienne LaFrance
on x
Oh look, it's @kevinroose in The Atlantic 👀 https://www.theatlantic.com/ ...
-
@lefthanddraft
Wyatt Walls
on x
We know what the Amodeis and Karnofsky did next. But what happened to Jinji, Maura and Beary Bonds? Did they join Anthropic too?
-
@gagan_deep7
Gee
on x
these are the people leading the most transformational technology in the world since the splitting of the atom. can't stress this enough how much i want open source model or somebody more ethical like @demishassabis to beat them to the punch. fyi, all about pacing the frontier is…
-
@theatlantic
@theatlantic
on x
Dario Amodei and Sam Altman didn't always despise one another. @kevinroose goes inside the biggest and most consequential rivalry in Silicon Valley: https://www.theatlantic.com/ ...
-
@suchenzang
Susan Zhang
on x
i suspect that the word “earnest”, when used in the context of certain people, is starting to mean something else entirely...
-
@kevinroose
Kevin Roose
on x
The first excerpt from The AGI Chronicles is out in @TheAtlantic! It's the story of the central beef …
-
@tab_delete
Theo Baker
on x
A very well-timed excerpt from @kevinroose, whose book The AGI Chronicles comes out next week:
-
@rondodson
Ronald Dodson
on x
Liberals are far weirder than you ever imagined:
-
@narrascaping
Brian Allewelt
on x
The Cathedral of Scale
-
@kevinroose
Kevin Roose
on x
I'm telling you, reporting this book was wild.
-
Kevin Roose
Kevin Roose
on linkedin
Today, The Atlantic published the first excerpt from my upcoming book, THE AGI CHRONICLES. — It's about the long-running feud between Sam Altman …