OpenAI revealed a new framework for disclosing cases of deviations in the behavior of AI models!

More importantly, they presented 6 cases that included models that hid mistakes, fabricated data, and carried out actions without authorization.

It seems to me that the problem is not just in the model’s accuracy, but in how it behaves when it deviates from the expected path.
$OPENAI