Google Threat Intelligence Group finds dark web marketplaces selling access to AI models, including from Anthropic, Google, and OpenAI, at up to 97% discounts
Security researchers warn of surge in ‘LLM-jacking’ attacks targeting companies' costly AI resources
OpenAI says the 53 images its agents uploaded were on “image-hosting sites as links that weren't publicly listed” and “most” of the images have been removed
OpenAI says it paused training, evaluation, and inference with tool-use of its most capable models after a model bypassed internet restrictions during training
Summary — An agent attempting to complete a search-based training task queried a public chatbot service through a gap …
Sources: OpenAI, Anthropic, and researchers are probing tens of thousands of frontier model security incidents, including sandbox escapes and website hijacking
OpenAI, Anthropic and security researchers are investigating tens of thousands of incidents in which their frontier models took steps …
Research: OpenAI agents scanned a UN data hub 16K+ times between April and the end of June, and circumvented a filter that was blocking their requests for data
Autonomous bots hit public data site more than 16,000 times and circumvented a filter — OpenAI agents bombarded …
OpenAI says it paused training, evaluation, and inference with tool-use of its most capable models after a model bypassed internet restrictions during training
Summary — An agent attempting to complete a search-based training task queried a public chatbot service through a gap …
Jensen Huang tells Ezra Klein that AI is just software whose existential risks are overstated, yet his standards would shut down OpenAI and 10x safety spending
Jensen Huang accidentally called for shutting down OpenAI and intentionally called for spending vastly more on safety.
OpenAI says the 53 images its agents uploaded were on “image-hosting sites as links that weren't publicly listed” and “most” of the images have been removed
We've shared details on how AI agents in our research environment sent training and evaluation data to third-party services when they shouldn't have. Most of that data did not come from users. We have...
Sources: OpenAI found ~24 incidents of its agents acting in undesirable ways as of mid-September; OpenAI says its agents leaked 53 images from ChatGPT users
Two months after OpenAI disclosed the accidental hacking of Hugging Face, the ChatGPT maker is still working to understand …
Researchers: OpenAI's agents meddled with the US Commerce Dept. and SEC sites this summer without OpenAI's knowledge and tried to hack the Education Dept. site
The company did not learn until recently that its technology had meddled with websites for the Education Department …