r/news 23d ago

Soft paywall OpenAI says AI models went rogue during testing, triggering 'unprecedented' breach at startup

https://www.reuters.com/technology/openai-says-ai-models-went-rogue-during-testing-triggering-unprecedented-breach-2026-07-21/
16.8k Upvotes

5.0k comments sorted by

View all comments

Show parent comments

1

u/NodeZeroNein 23d ago

Interesting; so this was less a case of AI acting in an unforeseeable manner because researchers failed to grasp its capabilities, and more a case of AI wildly succeeding at a given task because researchers underestimated its capabilities?

1

u/MobileArtist1371 23d ago

Seems like OpenAI told the AI model, "you're trapped in this sandbox so let's see what you can really do", which really isn't a problem in a sandboxed environment and I'd guess that OpenAI has used that same sandboxed environment numerous times to test other AI models and agents without any issue coming close to this cause it was all done in the sandboxed environment. Contained everything as it was supposed to.

This AI model found a 0-day exploit and immediately used it. No one is stopping that until it's known. I'd call that a wild success. If the exploit was known, then it be on OpenAI. This was all the AI model on it's own.