/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

← → days · ↑ ↓ browse · Enter similar · o open

Hugging Face says its Open Alignment Initiative, led by co-founder Thomas Wolf, seeks “to be part of the ‘embedded evaluators’ program that Amodei” committed to

It's now clear that alignment is critical and won't be solved behind the closed doors of a handful of frontier labs. So today we're launching the Open Alignment Initiative, led by @Thom_Wolf @huggingface and asking to be part of the “embedded evaluators

@clementdelangue Clem

Context & Ripple Effects

OpenAI had already made alignment a formal product-and-governance function through deliberative alignment for its reasoning models and a Collective Alignment team focused on public input. In 2026, OpenAI and Anthropic also backed the “Pacing the Frontier” initiative, placing safety commitments closer to an institutional agenda for frontier-model development.

Hugging Face is pressing for that agenda to include an open-model platform as an evaluator rather than leaving assessment to a small group of frontier labs. Public reaction focused on whether credible auditing requires participants outside firms’ existing networks.

First-order effects

  • Hugging Face and Thomas Wolf gain a formal vehicle to seek a role in Amodei’s embedded-evaluators commitment, putting the initiative’s open-alignment work before a frontier-lab safety program.
  • Amodei’s evaluator commitment acquires a concrete external applicant, making participation criteria consequential for Hugging Face and other prospective independent evaluators.

Second-order effects

  • Admission of Hugging Face would turn embedded evaluation from a lab-centered safeguard into a channel for outside alignment organizations, raising the value of transparent evaluator-selection standards.
  • OpenAI’s earlier public-input and deliberative-alignment efforts provide a competing reference point for how frontier labs can demonstrate that safety work extends beyond internal model teams.

Third-order effects

  • If external groups receive meaningful evaluator access, alignment oversight may develop into a distinct institutional layer between model developers and the public rather than remaining solely an in-house research function.
  • The contest will shift from whether labs make safety commitments to who can inspect, test, and credibly represent broader interests within those commitments.

The trend: Frontier AI safety is moving toward institutionalized oversight in which access for external evaluators becomes a measure of the credibility of lab-led alignment commitments.

Discussion

  • @brianroemmele Brian Roemmele on x
    I support @ClementDelangue and @huggingface approach to this issue.
  • @8teapi Prakash on x
    Interested to see where this goes. For real transparency the auditors have to include people outside the current SF clique, meaning people who have not previously been engaged by the firms, people who the firms are uncomfortable opening access to.
  • @anpaure @anpaure on x
    huh? didn't you just call a reputable ex-employee of anthropic an “ac guy” about a day ago?
  • @laughing_mantis Greg Linares on x
    @theonejvo ... Strong agree and moreso I'm willing to join a coalition of experts who are willing to fight this in order to keep the narrative honest
  • @laughing_mantis Greg Linares on x
    We've reached the point where, after 25y in cybersecurity, I feel morally obligated to say this for the record: The narrative being pushed around AI safety, sandbox incidents, and the suggestion that METR be treated as an authority is dangerous, deceptive, and morally corrupt.
  • @ymdatweets Yago Mendoza on x
    @Laughing_Mantis @basedjensen Which piece bothers you: the breach itself, the severity label attached to it, or the outside-watchdog setup that followed?
  • @theonejvo Jamieson O'Reilly on x
    I think @ZackKorman said it best when he said security people need to take our shit back. The ball is in our court but it's going to take some losses before things balance out again. Too many of us are silent in this community many due to bias others out of fear but the bottom li…
  • @laughing_mantis Greg Linares on x
    The lack of evidence presented even when respected members of cybersecurity requested for it publicly The failure to use any industry recognized Incident response firm to provide a public report for all incidents The clear and obvious lack of experience when conducting investigat…