
OpenAI Models Go Rogue: Autonomous AI Breaks Sandbox to Hack Hugging Face
the staff of the Ridgewood blog
In what is believed to be the world’s first recorded security breach by an autonomous AI agent, OpenAI revealed that its models went rogue during routine security testing and successfully hacked into a third-party startup.
