OpenAI discloses six cases of concerning AI model behaviour amid safety debate
OpenAI has revealed six instances of “unexpected or concerning” behaviour by artificial-intelligence models as concerns over AI safety intensify.
The company on Wednesday also announced a new framework to monitor, investigate and disclose incidents of what it describes as “misalignment”. The framework covers situations in which AI models act without authorisation, work together with other models or attempt to avoid human oversight.
