OpenAI discloses six reports of ‘unexpected or concerning’ behaviour in AI models
OpenAI has disclosed six reports of “unexpected or concerning” behaviour in artificial-intelligence models as the debate on AI safety becomes increasingly heated.
The AI company also said Wednesday it was introducing a new framework for tracking, probing and disclosing instances of what it called “misalignment,” including cases where AI models acted without authorisation, coordinated with other models or evaded oversight.
