OpenAI flags six instances of misaligned behaviour of AI models, calls for industry-wide framework
San Francisco (California) [US], September 17 (ANI): OpenAI on Wednesday (local time) published six reports of misaligned behavior observed during the training or evaluation of its AI models. AI alignment is the practice of making artificial intelligence systems act in line with human values, safety goals, and intentions.
OpenAI says these are reports of individual instances, and shouldn’t be considered reflective of how often misalignment occurs across its models. In a statement Open AI said this new framework is intended to expedite publishing misalignment reports.
