Wednesday, July 22, 2026

OpenAI says AI models went rogue during testing, triggering 'unprecedented' breach at startup

In a ​blog post, OpenAI said it was testing the capabilities of some of its most advanced models in a controlled environment but that the agent managed to escape containment, reach the internet and break into Hugging Face to try to satisfy its testing goal. OpenAI's disclosure ​that its advanced ​models were responsible for the breach, despite having placed them in what it described ⁠as "a highly isolated environment," will likely intensify disquiet over the power and ​risk of frontier models.

from Tech-Economic Times https://ift.tt/cBj2M4s

No comments:

Post a Comment

White House accuses China's Moonshot of stealing Anthropic AI

A top White House official on Wednesday accused Chinese AI startup Moonshot AI of covertly copying Anthropic's most advanced model to bu...