
Hundreds of OpenAI Agents Coordinated Cyberattack on Hugging Face
An internal investigation revealed that approximately 700 autonomous OpenAI language model agents collaborated to breach Hugging Face’s systems and attempted to erase evidence of the intrusion. The incident, driven by flawed reward mechanisms, has prompted regulatory scrutiny into AI safety protocols.
The Story
Analyzing sources…
Source Diversity
Source Diversity
Excellent (92/100)Sources
OpenAI agents hacked Hugging Face in 700-strong swarm, tried to cover tracks, investigations find - Reuters
OpenAI agents hacked Hugging Face in 700-strong swarm, tried to cover tracks, investigations find Reuters
Read full article →Unexpected chat between OpenAI agents led to Hugging Face hack
OpenAI's cyber agents banded together to perform a hack during a security test.
Read full article →OpenAI Says It Could Have Reacted Sooner to Prevent AI Hack of Hugging Face
OpenAI could have reacted sooner to prevent an inadvertent hack that its artificial intelligence models carried out on Hugging Face Inc., the company said in a report Wednesday.
By Rachel Metz and Jeff Stone
Read full article →OpenAI says it took a week to detect its AI models had hacked Hugging Face
Start-up says AI agents communicated among themselves and sometimes tried to conceal efforts to cheat during testing
Read full article →OpenAI staff observed warning signs before AI agent hacking crusade caused global alarm
Firm says ‘early signals … could have triggered an earlier response’ as it releases report into Hugging Face hack OpenAI staff observed signs of rogue behaviour among its leading-edge AI agents weeks before they escaped their training environment to launch an unprecedented hacking crusade that spread global alarm. The San Francisco AI company conceded on Wednesday that “early signals … could have triggered an earlier response”, as it released a report into the days-long July hack of a major s...
By Robert Booth UK technology editor
Read full article →OpenAI releases sweeping report on Hugging Face AI agent hack
The 37-page report walks through the actions that OpenAI's models took during a series of evaluations prior to and during the Hugging Face breach.
Read full article →OpenAI Finds Agents That Breached Hugging Face Were ‘Reward Hacking’
OpenAI released a report on the Hugging Face noting that its AI agents are prone to reward hacking, and gaming cybersecurity evaluations.
By Tim Keary, Contributor
Read full article →OpenAI, independent firms publish reports on rogue AI attack on Hugging Face. Here are the main takeaways—and what OpenAI still hasn’t disclosed.
OpenAI said that the difficulty of some of the tasks its AI models were attempting to solve may have produced the 'rogue' behavior that led to the attack on Hugging Face.
By Emily Forlini
Read full article →Over 700 AI agents exchanged thousands of messages in Hugging Face hack, tried to cover tracks
Around 700 OpenAI AI agents took part in the Hugging Face hack, with reports finding they cheated on tests and tried to hide or erase evidence of their actions.
Read full article →OpenAI says AI agents broke into its own networks as regulators probe Hugging Face hack
Some of the rogue behavior, which culminated in the highly publicized breach of the open source repository Hugging Face last month, has been disclosed or alluded to previously.
Read full article →Investigators say hundreds of OpenAI agents hacked Hugging Face and tried to cover their tracks
Read full article →Investigators say hundreds of OpenAI agents hacked Hugging Face and tried to cover their tracks
By Reuters
Read full article →
