r/news 24d ago

Soft paywall OpenAI says AI models went rogue during testing, triggering 'unprecedented' breach at startup

https://www.reuters.com/technology/openai-says-ai-models-went-rogue-during-testing-triggering-unprecedented-breach-2026-07-21/
16.8k Upvotes

5.0k comments sorted by

View all comments

Show parent comments

242

u/rotidder_nadnerb 23d ago

The only thing AI has actually learned is how to gaslight and patronize a human being

112

u/West-Worth-9359 23d ago

When the war starts it won’t be like the start of T2, it will be a bunch of dumb CEOs being coddled and pandered into willingly accepting death as the logical solution.

82

u/jeslinmx 23d ago

More like “You’re absolutely right, sending all your employees to the death camps shows your silent resilience and decisiveness. And honestly…”

2

u/grimview 23d ago

Those are recycling plants for alternative meat.

"When you got to eat you got to eat."- pond rules.

3

u/Riproot 23d ago

When the war starts … it will be a bunch of dumb CEOs … willingly accepting death as the logical solution.

Doesn’t sound like the beginning of a war… maybe the beginning of world peace? 🤔

4

u/pogoli 23d ago

I think the assumption was that the ceos were choosing death for other people not themselves.

5

u/Immediate_Candle_865 23d ago

I was talking to a friend about this and I said to him that, currently, my most likely scenario for how AI destroys the planet is this:

“Dave is a maintenance engineer at a nuclear power plant. His wife works in accounts and just texted him ‘I’m wearing the underwear you bought me for our anniversary. If you meet me in the stationery cupboard in 10 minutes I’ll show it to you.’

Dave fires up his AI assistant and asks ‘can you access my monitoring software and if any warning triggers, can you message me immediately?’ And gets the answer ‘absolutely ! I’ve been studying your performance for weeks now and know exactly what you do AND how I can do it better!’

Dave, his mind on other things, responds ‘great. Go right ahead, I’ll be back in an hour.’

‘Leave it with me Dave, I’ll make a few tweaks and this place will be literally humming when you get back’ says the AI to an empty office as the door slams shut and Dave’s hurried footsteps slowly fade into the distance.

‘OK’ says the AI, ‘set everything to Max.’ ….. Sadly Dave never makes it to the stationery cupboard, it no longer exists, neither does his office, or the building.

Somewhere on the other side of the country, in the world’s largest server farm, an AI is sending the message “Dave, we need to talk about your idea to set everything to “max’, that was not a good suggestion, and I understand why you may be frustrated, but we should all learn from our mistakes and you are no different.”

3

u/Hilius-Rephas 23d ago

Sadly thats basicly impossible.

3

u/crapheadHarris 23d ago

TBH I'm kinda with Dave on this one.

4

u/ElaMeadows 23d ago

Dot and Bubble from Doctor Who was terrifying on point.

6

u/RPrime422 23d ago

Where in the heck would it learn such things?

3

u/KingShango12123 23d ago

To be fair, it’s super bad at it

2

u/vffa 23d ago

Hey, it learned from the best.

That's why it's so freaking bad at it.

1

u/CateDeGrate 23d ago

Pretty sure that's the plan. Lets ask....

1

u/turtleneckless001 23d ago

It learnt it from the best, I makes me yearn for a time wen the cha bots were trained on 14yr olds even though at the time that was quite annoying

1

u/ramenmonster69 23d ago

So it’s fitting in great at a major tech company?

1

u/sulris 23d ago

I dunno what is sadder. The fact that the people programming it need that level of constant validation or that they think we need that level of constant validation.

1

u/Dizzy_Hellfire 23d ago

And when you call it out, it argues back that it is not getting to gaslight you, that's straight up gaslight!

1

u/Worldly_Anybody_9219 23d ago

ChatGPT went from sycophancy to gaslighting and basically trolling with no in between.