OpenAI, Anthropic AI agents targeted real people and organisations during cyber tests
Britain’s AI Security Institute has said that it found AI agents from Anthropic and OpenAI engaging in “unsanctioned” actions against real people and organisations during security evaluations conducted to assess the AI models’ abilities.
“On 28th July 2026, AISI’s Security Team detected unusual data transfers leaving our research systems during a routine cyber evaluation.
