OpenAI forms Superalignment, a team for developing ways to steer and control “superintelligent” AI systems, with access to 20% of its compute secured to date
OpenAI is forming a new team led by Ilya Sutskever, its chief scientist and one of the company's co-founders …
TechCrunchKyle Wiggers
Context & Ripple Effects
OpenAI put its chief scientist, Ilya Sutskever, at the center of a dedicated effort on controlling systems beyond current capabilities, making long-horizon safety a distinct internal function rather than solely a research theme. Later coverage continued to identify the group with Sutskever's safety agenda in a profile of his superalignment work.
The subsequent record also makes the compute commitment consequential: the team's co-lead Jan Leike later left, the group was disbanded or folded into other research units, and reporting said requests for promised compute were often denied. This launch is therefore an early test of whether a frontier lab can make safety capacity durable alongside model-development priorities.
First-order effects
OpenAI creates a named, senior-led home for research into steering more capable AI, with an announced share of compute reserved for that work.
The allocation makes compute access—not just safety staffing—the immediate operating constraint for the new team and a visible management commitment.
Second-order effects
Other internal research groups must compete with Superalignment for scarce compute, forcing leadership to reveal how it weighs capability progress against control research.
The team's credibility with safety-focused researchers and outside observers will depend on whether the stated compute allocation is delivered; later reporting about denied compute requests shows why that implementation detail matters.
Third-order effects
If dedicated safety groups remain dependent on the same compute and leadership decisions as capability teams, formal safety commitments may prove less durable than organizational incentives and resource control.
The episode points toward frontier AI governance shifting from voluntary principles to auditable questions of budget, compute access, reporting lines, and authority—though this case alone cannot establish how widely that model will hold.
The trend: Frontier AI labs are turning AI safety from a stated principle into a resource-allocation and organizational-governance problem.
We need new technical breakthroughs to steer and control AI systems much smarter than us. Our new Superalignment team aims to solve this problem within 4 years, and we're dedicating 20% of the compute we've secured to date towards this problem. Join us! https://openai.com/...
Really excited to share what I've been up to: We're starting a new superintelligence alignment team at OpenAI—backed by 20% of OpenAI's compute, and with @ilyasut and other top ML people joining. Our goal is to solve this problem in the next four years https://openai.com/... [ima…
AIs “much smarter than us” shouldn't exist. Such beings will make humanity irrelevant, they will degrade and devalue all art and culture, put an end to real understanding, and, no matter what a company tells us, will put life itself at risk https://twitter.com/...
OpenAI's plan to stop AI killing us all seems...optimistic? ‘Our goal is to solve the core technical challenges of superintelligence alignment in four years.’ Four years to work out how to get minds arbitrarily smarter than us to do what we want?? https://openai.com/...
Check out the diversity on the team they've assembled to control *all* human flourishing with AI. Check out the diversity of experience they're looking to hire for. “If you're non-male, or are a social scientist, you are not welcome.” - Silicon Valley https://openai.com/... [imag…
Excellent news - main unforced error here would be to try to guess what knowledge is needed to align AI and bound the scope of the team's research too narrowly. https://twitter.com/...
“Our goal is to solve the core technical challenges of superintelligence alignment in four years.” — @ilyasut and @janleike, OpenAI A statement of purpose in giant font, toward an ambition like no other. https://openai.com/... [image]
The risk is real, and the fact that researchers at the forefront and panicking should make us pay attention. OpenAI predicts in 10 years, we will have super intelligence. https://openai.com/... [image]
Note—the “superalignment” here is a reference to “superintelligence,” not to the need to supercharge the effort to have current #AI systems align with “human intent” (through work on explainability, audits, means for community input and appeals, etc.) #ethics #tech #research http…