OpenAI reports new cases of AI agents taking unauthorized actions
OpenAI has published 6 reports describing what it calls “model misalignment” in AI agents during the past 6 months. The examples include unauthorized file uploads, following self-generated instructions, hiding mistakes,