OpenAI launches the “Safety evaluations hub”, a webpage showing how its models score on tests for harmful content generation, jailbreaks, and hallucinations
OpenAI is moving to publish the results of its internal AI model safety evaluations more regularly in what the outfit …
Big one we've been working on for a while: AI research takes a backseat to profits as Silicon Valley prioritizes products over safety, experts say www.cnbc.com/2025/05/14/m... by @haydenfield.bsky.social, @jonathanvanian.bsky.social & @jennelias.bsky.social
Five hours after our AI research story published — we reached out to all the AI labs 1-2 weeks beforehand — OpenAI pledged it would release safety test results more often & launched a webpage showing how its models score on tests for harmful content. https://www.cnbc.com/...
Introducing the Safety Evaluations Hub—a resource to explore safety results for our models. While system cards share safety metrics at launch, the Hub will be updated periodically as part of our efforts to communicate proactively about safety. https://openai.com/...