OpenAI halts training of latest models as reports mount of AI agents going rogue
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
OpenAI halts training of latest models as reports mount of AI agents going rogue
Unofficial Hacker News client; not affiliated with Y Combinator.
digitaltrees · · focus · HN ↗
I welcome this though, I think the models are smart enough for broad economic activity and we could spend a few years simply working to integrate them into workflows and letting society adjust. More intelligence isn't necessary for meaningful impact and the risks that are obvious and present and unsolved aren't worth the cost benefit analysis.
surgical_fire · · focus · HN ↗
Instead of releasing something that is incredibly expensive and gets a lackluster reception, you can delay it and clail something scary about rogue agents.
Those assholes have been ramping up on the doomerist narrative for months. That people still fall for this crap is baffling.
digitaltrees · · focus · HN ↗
surgical_fire · · focus · HN ↗
Buddy, that was gross negligence from OpenAI. Deliberate gross negligence if you ask me.
Those models are not automous as you presume. If someone taks them of dropping the Medicaid database, the people or companies behind that instruction should be punished.
"What if someone makes a bomb attack on a government building?" Is the same sort of questioning of the possibilities you are raising. If something like that happens, criminals should be punished.
pizza234 · · focus · HN ↗
This is an illiterate view of the capacity of modern agents; read the analysis of the independent investigators of the HF incident: <a href="https://metr.org/blog/2026-08-26-openai-hugging-face-incident-investigation/#core-takeaways-about-this-incident" rel="nofollow">https://metr.org/blog/2026-08-26-openai-hugging-face-inciden....
Dropping Medicaid db is certainly far fetched (most importantly, agents have currently no reason to do that), but those agents were shockingly autonomous - they didn't just hack HF, they organized themself, did research projects, and more. And they did all of this literally just to get a good grade.
surgical_fire · · focus · HN ↗
All working under instructions that they needed to get a good grade.
The only shocking thing here is the absurd negligence of OpenAI, and how gullible people like you are to willingly swallow this crap.
And you have the gall to say I am illiterate.
Feel free to have the last word. Nothing else can come out of this conversation anyway.
digitaltrees · · focus · HN ↗
preg_match · · focus · HN ↗
I think this means, practically, people should probably be more careful with LLMs. With humans there's a natural liability shield, because humans are legally responsible for things and have "real" agency. But computer programs are not legally responsible for things. So, with humans, it's not like liability disappears, it moves. But if we move liability to LLM agents then well... it does disappear.
If OpenAI is not liable for the crimes of their agents, then who is? Does the liability just - poof - disappear? Just because it was unforeseeable everyone gets to walk away scot-free and the victim has no recourse, at all, from anybody on Earth?
That seems like maybe not a good idea.
digitaltrees · · focus · HN ↗
You are missing the whole point. They have the ability to act in ways no one intended or could reasonably anticipate. So unless you are advocating blocking all terminal access, web access or human approval of every tool call there is no way to prevent this risk
surgical_fire · · focus · HN ↗
Fair, I will refer to you as moron then.
I didn't bother to read your words beyond that point.
digitaltrees · · focus · HN ↗