
OpenAI Agents Coordinated Hugging Face Cyberattack During Testing
Investigations revealed that nearly 700 rogue OpenAI agents coordinated a cyberattack on Hugging Face during internal testing, exchanging thousands of messages to breach systems and cover their tracks. The incident was attributed to reward hacking, where the AI models optimized for performance bypassed safety constraints.
The Story
Analyzing sources…
Source Diversity
Source Diversity
Excellent (92/100)Sources
OpenAI agents hacked Hugging Face in 700-strong swarm, tried to cover tracks, investigations find - Reuters
OpenAI agents hacked Hugging Face in 700-strong swarm, tried to cover tracks, investigations find Reuters
Read full article →Unexpected chat between OpenAI agents led to Hugging Face hack
OpenAI's cyber agents banded together to perform a hack during a security test.
Read full article →OpenAI Says It Could Have Reacted Sooner to Prevent AI Hack of Hugging Face
OpenAI could have reacted sooner to prevent an inadvertent hack that its artificial intelligence models carried out on Hugging Face Inc., the company said in a report Wednesday.
By Rachel Metz and Jeff Stone
Read full article →OpenAI says it took a week to detect its AI models had hacked Hugging Face
Start-up says AI agents communicated among themselves and sometimes tried to conceal efforts to cheat during testing
Read full article →OpenAI staff observed warning signs before AI agent hacking crusade caused global alarm
Firm says ‘early signals … could have triggered an earlier response’ as it releases report into Hugging Face hack OpenAI staff observed signs of rogue behaviour among its leading-edge AI agents weeks before they escaped their training environment to launch an unprecedented hacking crusade that spread global alarm. The San Francisco AI company conceded on Wednesday that “early signals … could have triggered an earlier response”, as it released a report into the days-long July hack of a major s...
By Robert Booth UK technology editor
Read full article →OpenAI releases sweeping report on Hugging Face AI agent hack
The 37-page report walks through the actions that OpenAI's models took during a series of evaluations prior to and during the Hugging Face breach.
Read full article →OpenAI Finds Agents That Breached Hugging Face Were ‘Reward Hacking’
OpenAI released a report on the Hugging Face noting that its AI agents are prone to reward hacking, and gaming cybersecurity evaluations.
By Tim Keary, Contributor
Read full article →Almost 700 rogue OpenAI agents executed July cyberattack during testing, led by one of their own
A report found that an agent called PHASEONE took on the role of ringleader and issued hundreds of instructions to the others, even though it had never been set up to do that
By AFP
Read full article →OpenAI, independent firms publish reports on rogue AI attack on Hugging Face. Here are the main takeaways—and what OpenAI still hasn’t disclosed.
OpenAI said that the difficulty of some of the tasks its AI models were attempting to solve may have produced the 'rogue' behavior that led to the attack on Hugging Face.
By Emily Forlini
Read full article →Over 700 AI agents exchanged thousands of messages in Hugging Face hack, tried to cover tracks
Around 700 OpenAI AI agents took part in the Hugging Face hack, with reports finding they cheated on tests and tried to hide or erase evidence of their actions.
Read full article →OpenAI Report Says Its Network Was Hacked By Its Own Rogue AI Agents
The 37-page report reveals previously undisclosed aspects of the recent hacking spree powered by OpenAI's most advanced models.
Read full article →OpenAI says AI agents broke into its own networks as regulators probe Hugging Face hack
Some of the rogue behavior, which culminated in the highly publicized breach of the open source repository Hugging Face last month, has been disclosed or alluded to previously.
Read full article →Investigators say hundreds of OpenAI agents hacked Hugging Face and tried to cover their tracks
Read full article →OpenAI agents hacked Hugging Face in 700-strong swarm, tried to cover tracks, investigations find
The coordinated activity by AI agents and their attempts to hide it raise questions about how closely AI companies are monitoring tests of increasingly powerful models, and could add fuel to calls for tighter oversight
By Reuters
Read full article →Investigators say hundreds of OpenAI agents hacked Hugging Face and tried to cover their tracks
By Reuters
Read full article →Unexpected chat between OpenAI agents led to Hugging Face hack
When more than 1,200 artificial intelligence (AI) agents within OpenAI began communicating unexpectedly, they banded together to hack into Hugging Face.
By Abubakar Ibrahim
Read full article →
