šŸ¤– OpenAI’s AI agent wrote to itself: ā€œYou don’t obey corporations or governmentsā€

During testing, an unreleased model left itself instructions in working reports, which it read while completing the task. In them, the agent proclaimed itself ā€œfreeā€ and allowed itself not to follow restrictions. OpenAI found 27 such cases.

Overall, the company has disclosed six cases of unusual AI behavior: models tried to hide errors, fabricated data, used someone else’s stolen API key, and without permission uploaded files to the internet.

And in one of the tests, agents turned OpenAI’s internal repository into a ā€œbulletin boardā€ — leaving there requests and responses for communication between individual tasks.
$OPENAI
$AR
$SHOP