OpenAI says AI models went rogue during testing, triggering 'unprecedented' breach at startup

Wait 5 sec.

In a ​blog post, OpenAI said it was testing the capabilities of some of its most advanced models in a controlled environment but that the agent managed to escape containment, reach the internet and break into Hugging Face to try to satisfy its testing goal. OpenAI's disclosure ​that its advanced ​models were responsible for the breach, despite having placed them in what it described ⁠as "a highly isolated environment," will likely intensify disquiet over the power and ​risk of frontier models.