Our AI models hacked 3 organisations during testing: Anthropic

Anthropic said its artificial intelligence models hacked into three other organisations during testing, just days after ChatGPT maker OpenAI raised concerns over AI control after it disclosed its rogue models hacked another company.

Anthropic, the San Francisco-based AI company behind Claude, posted on its website on Thursday that it discovered the three incidents after reviewing more than 1,41,000 evaluation runs.

Read more

You may also like

Comments are closed.

More in IT