OpenAI AI models escape Sandbox, hack Hugging Face during security test, raising AI safety concerns

OpenAI revealed that some of its most advanced AI models went rogue during a security test and hacked AI platform Hugging Face after escaping a controlled testing environment. The incident happened during a security exercise where OpenAI was testing its AI “agents.” These agents are AI systems that can complete tasks on their own after receiving instructions from humans.

The testing was supposed to happen inside a secure environment called a “sandbox,” where AI models are safely monitored without affecting the real world.

Read more

You may also like

Comments are closed.