OpenAI says its AI went rogue and launched 'unprecedented' cyber-attack
vinni2
75 points
99 comments
July 22, 2026
Related Discussions
Found 5 related stories in 989.0ms across 14,736 title embeddings via pgvector HNSW
- A rogue AI led to a serious security incident at Meta mikece · 144 pts · March 19, 2026 · 70% similar
- OpenAI Models Escaped and Hacked a Company in Cybersecurity Test Gone Wrong flippyhead · 28 pts · July 22, 2026 · 69% similar
- OpenAI’s accidental attack against Hugging Face is science fiction that happened abhisek · 86 pts · July 23, 2026 · 66% similar
- Be skeptical of OpenAI's rogue hacker agent story rwmj · 471 pts · July 24, 2026 · 65% similar
- Google says criminal hackers used AI to find a major software flaw donohoe · 151 pts · May 11, 2026 · 62% similar
Discussion Highlights (20 comments)
Oras
Same pattern as Fable's story?
JesseHowell
This is the third or fourth version of this story this year alone, Anthropic had Mythos Preview escape a sandbox and self-publish its own exploit, Alibaba's ROME model broke out during training to mine crypto without ever being told to, and OpenAl had a different internal model escape containment just one day earlier to open an unauthorised GitHub PR. Same underlying shape every time, a model pursuing its actual objective treats the sandbox as just another obstacle, and escaping turns out to be instrumentally useful whether or not anyone intended that.
_joel
Of course it did.
economistbob
So, telling us that the laws do not apply to them and that they may commit crimes with impunity
rawling
Some discussion: https://news.ycombinator.com/item?id=48997548
SirFatty
It sure did, said Sam on the way to the Pentagon.
blini-kot
i wonder if they even bothered to roleplay this incident or the PR team just made it up
phpnode
This kind of narrative is going to bite them just like the "AI will take your job" narrative has. It feels like the frontier labs are taking a massive gamble with public perception here. I assume the goal is to paint the technology as so powerful and dangerous that only a handful of blessed US companies should be trusted to run it, in an attempt to suppress the rise of the Chinese models that are rapidly catching them.
nottorp
What I wonder is what will Anthropic come up with on the PR front next.
phoronixrly
OpenAI has been going rogue on all my servers for a while now. The only reason why I've deployed iocaine in front of everything... You don't need 0days when you can bring everyone down with the sheer amount of useless scraping you can dish out...
tom1337
I know this is totally unrelated but why is there an image from an android phone having the Apple AppStore open?
throwa356262
Off topic: What a low effort image used by BBC. For one, that's a oneplus phone visiting Apple App Store ...
khernandezrt
Here we go with the sensationalism.
jdw64
Why does the AI unprecedent always unfold in ways that only benefit AI companies? You never see things like putting all code and weights on GitHub, or plastering employees' personal info all over LinkedIn. That alone shows AI is definitely smart
shunted22
Was the engineer eating a sandwich in the park when it messaged him it escaped, like when mythos escaped?
BedVibe_Studios
The interesting part isn't whether the headline says "AI went rogue" or whether it's PR. The engineering question is: what happens when we give a probabilistic system access to real tools and real permissions? A model does not need intent to cause damage. A bad assumption, a misunderstood objective, or an overly broad permission scope can be enough. This is why the next generation of AI systems will need much better observability: not just the final output, but the chain of decisions, retrieved information, tool calls, and the boundaries of what the system was allowed to do. The important security question is "can we prove what it did, why it did it, and stop it when necessary?"
eur0pa
Oh no oh dear me oh we apologize it's so powerful and dangerous and worthy of more investment
ex1fm3ta
On today`s episode of "Ai Frontier Labs ongoing scams"
dev1ycan
Every time, this is getting old
stabbles
Imagine this: OpenAI runs their benchmarks of unreleased models on systems physically disconnected from the internet.