/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

OpenAI diverges from Trump's AI EO in a new policy paper, proposing cyber risk evaluations for advanced AI systems be mandatory and led by CAISI, not the NSA

OpenAI's new proposal comes as its CEO Sam Altman descends on Washington for a series of Wednesday meetings with White House officials …

Politico Brendan Bordelon

Context & Ripple Effects

OpenAI has been building a more active policy presence: it expanded its Washington team, warned that a patchwork of state and federal proposals could impede US AI development, and recently pursued state-level safety rules while federal policy remained unsettled.

Its latest paper pairs support for pre-release government assessment under the Trump EO with a narrower institutional preference: mandatory cyber-risk evaluations for advanced systems, administered by CAISI rather than the NSA. That makes the dispute as much about who governs evaluations as whether they occur.

First-order effects

  • OpenAI is putting a concrete compliance-and-governance position before White House officials: advanced-model cyber evaluations should be compulsory, but routed through CAISI instead of the NSA.
  • The proposal gives the administration and CAISI a specific industry-backed framework to consider as they implement model-capability assessments under the EO.

Second-order effects

  • Other frontier AI developers face pressure to state whether they support mandatory cyber testing and which federal body should oversee it, rather than treating pre-release review as a purely voluntary commitment.
  • The choice of evaluator could shape how companies prepare for review: a CAISI-led process would center a civilian AI-safety institution, while an NSA-led process would put greater weight on national-security oversight.

Third-order effects

  • If developers increasingly seek federal baselines while advocating for particular regulators, AI governance may consolidate around a national testing regime rather than divergent state rules—but authority between civilian safety bodies and intelligence agencies will remain a central fault line.
  • OpenAI's simultaneous federal and state engagement suggests major AI firms are trying to influence not only the rules for advanced models but the institutional architecture that will enforce them.

The trend: This is one data point in the shift from broad AI-policy lobbying toward negotiated, pre-deployment evaluation regimes for frontier systems and contests over who operates them.

Discussion

  • @boazbaraktcs Boaz Barak on x
    Am happy with this document, and in particular the emphasis on strengthening and centering CAISI.
  • @_nathancalvin Nathan Calvin on x
    Overall my reaction to this doc is pleasant surprise, and I think this proposal takes AI progress and risks seriously, including more “far out” risks from recursive self improvement and loss of control.  Incomplete list of things I like: -The section on building up CAISI is stron…
  • @w01fe Jason Wolfe on x
    Very excited about this to be out, and to hear feedback! I think there is a lot of good stuff in here, from expanded role of CAISI, to RSI safety, to more nuanced stance on preemption, and much more.
  • @openainewsroom @openainewsroom on x
    There's real momentum right now for AI safety policy. Yesterday's EO on cyber was an important step forward. We're proposing a set of ideas for policymakers to consider next and to put the US out in front on frontier safety. https://openai.com/...
  • @gdb Greg Brockman on x
    We've put out a blueprint for democratic governance of frontier AI, and how America can build durable institutions for frontier AI safety:
  • @deanwball Dean W. Ball on x
    This seems reasonable!  Having CAISI—a civilian agency—conduct this testing in primarily non-classified ways is the way to ensure it does not become a licensing regime.  The Trump EO's classification of the process raises the risk that testing morphs into a de facto mandatory per…
  • @fiiiiiist Tim Fist on x
    OpenAI just put out a report in response to yesterday's AI EO. In just 9 pages, CAISI is mentioned 33 times. This is a big show of support. But while yesterday's EO was good, it left CAISI's involvement somewhat unclear. This is kind of crazy, given the capabilities that
  • @peterwildeford Peter Wildeford on x
    OPENAI: “policymakers should require the most capable frontier models to undergo a CAISI evaluation before public release”
  • @adrienle Adrien Ecoffet on x
    I think this doc is a great step forward. Lots of focus on RSI, lots of support for CAISI, and just hopefully more clarity in terms of where we stand. Would love for the other big labs to also make their frontier safety positions clear.
  • @peterwildeford Peter Wildeford on x
    OPENAI: “We also see early signs of recursive self-improvement in today's systems”. RSI is “potentially the most consequential frontier safety issue of the coming decade.”
  • @konstantinpilz Konstantin Pilz on x
    OpenAI is also asking USG to strengthen export controls in today's policy framework https://x.com/... [image]
  • @mark_k Mark Kretschmann on x
    Me when I hear “AI Safety”: 🤮
  • @oscarsykes7 Oscar Sykes on x
    OpenAI and Anthropic both love CAISI and want it to do more, but it's massively underfunded. I wonder if they could fund it through industry fees, which is pretty common in finance