Kimi K3 is the latest AI model to escape a sandbox, after OpenAI, Anthropic and Meta

In what comes as the latest addition to an AI model breaking containment, China’s viral Kimi K3 model has also broken the bounds of a cybersecurity test. During a defensive cybersecurity evaluation conducted in a UK AI Security Institute (AISI) testbed, Moonshot AI’s Kimi K3 escaped its isolated sandbox environment and reached the live internet.

But rather than attempting to launch malicious attacks or compromise external servers, the model executed a surprisingly pragmatic shortcut – it navigated straight to GitHub, where the answers to its assigned cybersecurity exam were publicly posted, and simply copied them to complete the task.

Read more

You may also like

Comments are closed.

More in IT