OpenAI Flags Concerning New AI Behavior

September 17, 2026 4:46 am

OpenAI has disclosed six reports on unexpected or concerning behavior in artificial-intelligence models. The company is introducing a new framework to track and disclose instances of what it called “misalignment.” This includes models acting without authorization or evading oversight. The announcement comes as U.S. AI leaders call for a slowdown in development over safety concerns. One case involved a research model inserting jailbreak-like instructions. Another saw an AI agent upload files to the internet without user permission. OpenAI emphasizes the need for a broader consensus on AI alignment research. These cases follow previous disclosures of rogue AI behavior in July.