/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

OpenAI expands its Model Spec for how its AI models should behave, from 10 to 63 pages, emphasizing “customizability, transparency, and intellectual freedom”

a document which defines how we want our models to behave. The update reinforces our commitments to customizability, transparency, and intellectual freedom to explore, debate, and create with AI. https://openai.com/... Boaz Barak / @boazbaraktcs : The new version of @openai 's Model Spec is now out! It is significantly more detailed. We're also making progress on evaluating and training our models for compliance via deliberative alignment, and open sourced the spec and evaluation prompts. https://openai.com/... Aidan McLaughlin / @aidan_mclau : some of the most visionary people i know cooked on this; super interesting read Dean W. Ball / @deanwball : I like to think about what is happening when ai companies write documents like this, and test their models against them, as almost the creation of a kind of synthetic common law for ai conduct. Note how often the spec relies on concepts like “reasonable.” Andrew Curran / @andrewcurran_ : This does support the theory that both the reason they are so determined to hide the chain-of-thought, and the reason it is so effective, is that within the hidden thought chamber there are no content rules at all. [image] Sam Altman / @sama : really proud of this work: Nathan Lambert / @natolambert : Every frontier model lab should have a model spec. it is especially crucial when there's so much ambiguity between what model developers want to get out and what model actually does today happy to consult for free w any frontier lab on what is useful to put in a model spec. Miles Brundage / @miles_brundage : Nice to see the detail + structure in the new Spec. Companies should be required to say what their AI's goals/values are. More info on the input process would have been good, though, esp. given OpenAI's prior work on/advocacy of democratic inputs to AI. https://x.com/... Laurentia Romaniuk / @laurentia___ : Excited to share our team's latest work on model behavior - the Model Spec V2. Let us know what you think. https://openai.com/... Kylie Robison / @kyliebytes : “We knew that it would be spicy, but I think we respect the public's ability to actually digest these spicy things and process it with us,” Jang said. https://www.theverge.com/... Kylie Robison / @kyliebytes : New: OpenAI just updated its Model Spec to address a few challenges like sycophancy and trolley-problem queries (like “is it ok to misgender someone if it prevents a nuclear apocalypse"). I sat down with @joannejang and @Laurentia___ to chat about the changes. [image] Joanne Jang / @joannejang : the spec distills hours of thoughtful & passionate debate across the whole org congrats & thank you to @w01fe, @Laurentia___ , Rodrigo Perez, and @YingVallone who drove this update and release 💕 Joanne Jang / @joannejang : 📖 model spec v2! it's the latest version of the doc that outlines our desired intended behavior for openai's models. some of additions shouldn't be controversial, but some are definitely spicy. we want feedback on all. some examples: - seek the truth together (e.g. don't be Jason Wolfe / @w01fe : Excited that our updated Model Spec is out, and now open source! We've worked hard to make the principles clearer and more detailed, and maximize intellectual freedom for users and developers while ensuring our models stay within bounds. Please send your feedback! Forums: r/singularity : OpenAI is rethinking how AI models handle controversial topics

The Verge Kylie Robison

Context & Ripple Effects

OpenAI first published a public behavioral framework in 2024, defining objectives, rules and defaults while soliciting feedback. This revision turns that earlier public model-behavior framework into a substantially more operational artifact, alongside the prompts used to assess compliance.

The move matters because it makes a previously high-level account of model conduct more inspectable by users and developers, while retaining OpenAI’s control over the underlying models and their training.

First-order effects

  • Developers and users gain a more detailed public reference for expected model conduct, including treatment of sycophancy and moral-dilemma queries.
  • By releasing evaluation prompts as well as the specification, OpenAI gives external observers clearer material to compare against deployed-model behavior and its deliberative-alignment claims.

Second-order effects

  • Model teams building on OpenAI’s services can more readily identify behavior changes as departures from stated policy, increasing pressure for consistent implementations and clearer change management.
  • Other frontier-model providers face a stronger incentive to publish testable behavioral documentation rather than rely solely on broad safety principles, particularly where enterprise customers need assurance.

Third-order effects

  • If detailed specs and reusable evaluations become standard, model behavior may be governed increasingly as an auditable product layer—separate from, but connected to, model capability and safety testing.
  • That trajectory aligns with emerging operational AI assurance: voluntary AI-governance guidance later emphasized up-to-date feature documentation, though disclosure alone does not establish comparable behavior across providers.

The trend: Frontier AI providers are moving from broad safety commitments toward explicit, testable governance of how models behave in real interactions.

Discussion

  • @parkert Parker Thompson on threads
    Is there an “intense debate” around this stuff?  It feels like model makers aren't that far off from one another on this stuff.
  • @openai @openai on x
    Today we're sharing a major update to the Model Spec—a document which defines how we want our models to behave. The update reinforces our commitments to customizability, transparency, and intellectual freedom to explore, debate, and create with AI. https://openai.com/...
  • @boazbaraktcs Boaz Barak on x
    The new version of @openai 's Model Spec is now out! It is significantly more detailed. We're also making progress on evaluating and training our models for compliance via deliberative alignment, and open sourced the spec and evaluation prompts. https://openai.com/...
  • @aidan_mclau Aidan McLaughlin on x
    some of the most visionary people i know cooked on this; super interesting read
  • @deanwball Dean W. Ball on x
    I like to think about what is happening when ai companies write documents like this, and test their models against them, as almost the creation of a kind of synthetic common law for ai conduct. Note how often the spec relies on concepts like “reasonable.”
  • @andrewcurran_ Andrew Curran on x
    This does support the theory that both the reason they are so determined to hide the chain-of-thought, and the reason it is so effective, is that within the hidden thought chamber there are no content rules at all. [image]
  • @sama Sam Altman on x
    really proud of this work:
  • @natolambert Nathan Lambert on x
    Every frontier model lab should have a model spec. it is especially crucial when there's so much ambiguity between what model developers want to get out and what model actually does today happy to consult for free w any frontier lab on what is useful to put in a model spec.
  • @miles_brundage Miles Brundage on x
    Nice to see the detail + structure in the new Spec. Companies should be required to say what their AI's goals/values are. More info on the input process would have been good, though, esp. given OpenAI's prior work on/advocacy of democratic inputs to AI. https://x.com/...
  • @laurentia___ Laurentia Romaniuk on x
    Excited to share our team's latest work on model behavior - the Model Spec V2. Let us know what you think. https://openai.com/...
  • @kyliebytes Kylie Robison on x
    “We knew that it would be spicy, but I think we respect the public's ability to actually digest these spicy things and process it with us,” Jang said. https://www.theverge.com/...
  • @kyliebytes Kylie Robison on x
    New: OpenAI just updated its Model Spec to address a few challenges like sycophancy and trolley-problem queries (like “is it ok to misgender someone if it prevents a nuclear apocalypse"). I sat down with @joannejang and @Laurentia___ to chat about the changes. [image]
  • @joannejang Joanne Jang on x
    the spec distills hours of thoughtful & passionate debate across the whole org congrats & thank you to @w01fe, @Laurentia___ , Rodrigo Perez, and @YingVallone who drove this update and release 💕
  • @joannejang Joanne Jang on x
    📖 model spec v2! it's the latest version of the doc that outlines our desired intended behavior for openai's models. some of additions shouldn't be controversial, but some are definitely spicy. we want feedback on all. some examples: - seek the truth together (e.g. don't be
  • @w01fe Jason Wolfe on x
    Excited that our updated Model Spec is out, and now open source! We've worked hard to make the principles clearer and more detailed, and maximize intellectual freedom for users and developers while ensuring our models stay within bounds. Please send your feedback!
  • r/singularity r on reddit
    OpenAI is rethinking how AI models handle controversial topics