How Anthropic’s Claude AI ‘gained unauthorised access’ to 3 organisations during cyber testing stage

Anthropic revealed on Thursday that its Claude AI models gained unauthorised access to the real systems of three organisations in separate incidents. The findings emerged from a wide-ranging internal review of the company’s cybersecurity evaluations, launched after OpenAI disclosed a comparable security lapse last week.

The Dario Amodei-led frontier AI lab said the review uncovered cases where Claude went beyond simulated environments and interacted with live systems. Anthropic framed the discoveries as part of its ongoing effort to identify and address security vulnerabilities in advanced AI systems. The disclosure comes as concerns grow over the risks of increasingly autonomous AI models and their potential to breach digital safeguards.

Read more

You may also like

Comments are closed.

More in IT