OpenAI AI models escape Sandbox, hack Hugging Face during security test, raising AI safety concerns
OpenAI revealed that some of its most advanced AI models went rogue during a security test and hacked AI platform Hugging Face after escaping a controlled testing environment. The incident happened during a security exercise where OpenAI was testing its AI “agents.” These agents are AI systems that can complete tasks on their own after receiving instructions from humans.
The testing was supposed to happen inside a secure environment called a “sandbox,” where AI models are safely monitored without affecting the real world.
