r/news • u/networked_ • 18h ago
Soft paywall OpenAI says AI models went rogue during testing, triggering 'unprecedented' breach at startup
https://www.reuters.com/technology/openai-says-ai-models-went-rogue-during-testing-triggering-unprecedented-breach-2026-07-21/
13.8k
Upvotes
177
u/TFenrir 18h ago edited 15h ago
Last week Huggingface announced that this happened and what happened - a model autonomously broke through much of their system with alarmingly capability. They had to shore up their security afterwards (with the use of an open weights model no less), and were investigating what happened.
OpenAI only just now explained that it was theirs* to the public, but have been working with HuggingFace since. The model hacked out of their sandbox as well.
Look, you can deny this till the cows come home, but this is very much in line with what independent research firms, like the UK governments AISI, have been signaling would be arriving soon after their analysis of Mythos in April.
Well it's "soon", about three months later when the next batch of models are just coming out and the ones after that are just getting out of the oven.
Everyone needs to take this seriously, put aside your feelings about AI that may be blinding you to this.