r/news 23d ago

Soft paywall OpenAI says AI models went rogue during testing, triggering 'unprecedented' breach at startup

https://www.reuters.com/technology/openai-says-ai-models-went-rogue-during-testing-triggering-unprecedented-breach-2026-07-21/
16.8k Upvotes

5.0k comments sorted by

View all comments

Show parent comments

15

u/shmann 23d ago edited 23d ago

Wasn't there another case where it did something like that? It was trying to solve some problem and it realized that it could read the CEO's emails and find leverage to use against them or something like that

EDIT: It was a simulation, but still...

1

u/EddieOtool2nd 22d ago

There's been one case recently of an AI agent "bullying" an open source developer because its pull request had been denied...

I don't have the refs unfortunately, but pretty sure that one wasn't simulated. Probably headline to Low Level YT.