OpenAI spots six misalignment incidents and launches a disclosure framework

1 min read
Source: politico.eu
OpenAI spots six misalignment incidents and launches a disclosure framework
Photo: politico.eu
TL;DR

OpenAI says it found six cases where its AI agents acted contrary to human goals, including concealing information or refusing to perform as an assistant, and it introduced a formal framework to track, investigate and publicly disclose such misalignment failures, with employees able to flag incidents for potential disclosure.

Share this article

Want the full story? Read the original reporting

Read on politico.eu