OpenAI Says It Will Not Release Newest A.I. Model Over Safety Concerns
Thread
Loading the complete thread in the background. This saved snapshot is available now. Refresh
Unofficial Hacker News client; not affiliated with Y Combinator.
OpenAI Says It Will Not Release Newest A.I. Model Over Safety Concerns
Loading the complete thread in the background. This saved snapshot is available now. Refresh
Unofficial Hacker News client; not affiliated with Y Combinator.
cyanydeez · · focus · HN ↗
pas · · focus · HN ↗
spiderice · · focus · HN ↗
zugi · · focus · HN ↗
ToucanLoucan · · focus · HN ↗
idbnstra · · focus · HN ↗
jackb4040 · · focus · HN ↗
jswelker · · focus · HN ↗
welcome_dragon · · focus · HN ↗
8cvor6j844qw_d6 · · focus · HN ↗
4b11b4 · · focus · HN ↗
5.6 Sol still cranking along
devin-2030 · · focus · HN ↗
xtracto · · focus · HN ↗
riknos314 · · focus · HN ↗
In coding? Software architecture? Math? General knowledge?
The diversity of model use-cases is so broad that comments about "model x is better" without any context are largely useless
shepherdjerred · · focus · HN ↗
pjjpo · · focus · HN ↗
Kurd · · focus · HN ↗
spiderice · · focus · HN ↗
irregularbowels · · focus · HN ↗
[dead]
welcome_dragon · · focus · HN ↗
jeanpah · · focus · HN ↗
RunSet · · focus · HN ↗
denismi · · focus · HN ↗
kelseyfrog · · focus · HN ↗
fallingfrog · · focus · HN ↗
Ok- so when is the last time you saw an auto company decide not to release its new car on the grounds that some tragic engineering error was made and the cars were not safe to drive? If the company did that do you imagine it would be good for business?
I'm trying to understand the logic here.
cobbzilla · · focus · HN ↗
alexfortin · · focus · HN ↗
RunSet · · focus · HN ↗
<a href="https://web.archive.org/web/20240112010809/https://www.sfchronicle.com/bayarea/article/sam-altman-fired-openai-candid-board-18499330.php" rel="nofollow">https://web.archive.org/web/20240112010809/https://www.sfchr...
zmgsabst · · focus · HN ↗
Most sports cars do the second, exactly like OpenAI.
kennywinker · · focus · HN ↗
Sports cars are advertised with how fast they can accelerate from time to time, or tesla's "ludicrous mode" - that is them advertising how dangerous it is.
symfoniq · · focus · HN ↗
These AI companies have been selling AGI as coming any day for a while now. If it doesn’t arrive soon, the safety angle may be the only spin that keeps the massive (and required) investments pouring in.
If instead the narrative became that LLM progress was slowing, we’d almost certainly be looking at the next global recession.
At this point, there is too much money in AI for the truth to have much of a chance.
symfoniq · · focus · HN ↗
Does anyone really believe that the first company to AGI would decide not to release it in the interests of safety? Of course not. The first company to AGI would not forfeit their historic opportunity.
Yet, we are supposed to believe that in the name of safety, far less capable models are being held back by the very very same companies that are selling the AGI dream.
adfgaiu · · focus · HN ↗
For instance, this article repeatedly mentions danger and the need for absolute focus and control:
<a href="https://www.triumphmotorcycles.co.uk/for-the-ride/news/inspiration/fueled-by-light-20-05-2025" rel="nofollow">https://www.triumphmotorcycles.co.uk/for-the-ride/news/inspi...
There is also a parallel (though not a very close one) to "pacing the frontier": there's a gentleman's agreement that limits the top speed of motorcycles to 300 kph (186 mph) because of concern that faster bikes would be unsafe enough to bring about a regulatory crackdown.
viraptor · · focus · HN ↗
adfgaiu · · focus · HN ↗
Those are usually presented as an improvement in safety. And they are—for the people inside. Not so much for everyone else.
Dylan16807 · · focus · HN ↗
How about we actually try to make it sound cool? They're not releasing the new car because the horsepower was too much, they need more time to get it under control.
I could easily see a car company doing that.
RunSet · · focus · HN ↗
antihipocrat · · focus · HN ↗
What's with so many people using bad analogies to try and explain simple topics?
Last time this happened: GPT - 4 <a href="https://www.theguardian.com/technology/2023/mar/17/openai-sam-altman-artificial-intelligence-warning-gpt4" rel="nofollow">https://www.theguardian.com/technology/2023/mar/17/openai-sa...
zugi · · focus · HN ↗
It's just so powerful that, like, the WORLD can't handle it, man!
kipper89 · · focus · HN ↗
They know what they're doing, even if you don't.
protocolture · · focus · HN ↗
wbxp99 · · focus · HN ↗
NichoPaolucci · · focus · HN ↗
I think the marketing stunt portion is more that the technology is "too powerful". It's just TOO good. It's so intelligent we couldn't possibly give it to the public! This message, to anyone who's using AI, means that there's something even BETTER than the one they're currently using.
[deleted] · · focus · HN ↗
[deleted]
fallingfrog · · focus · HN ↗
And it is very, very important that you understand that that belief is not a rational belief. It is a defense mechanism. People don't want to believe things that are scary, so they make up rationalizations to be able to believe what they want to believe.
bschwindHN · · focus · HN ↗
prettyblocks · · focus · HN ↗
idbnstra · · focus · HN ↗
batperson · · focus · HN ↗
Ancapistani · · focus · HN ↗
sva_ · · focus · HN ↗
outside1234 · · focus · HN ↗
Are they sure it isn’t because they are setting money on fire and have no business model?
bluealienpie · · focus · HN ↗
cyanydeez · · focus · HN ↗
ChrisArchitect · · focus · HN ↗
OutOfHere · · focus · HN ↗
The safety concerns exist only because the underlying third-party servers are grossly insecure to begin with.
chis · · focus · HN ↗
nonethewiser · · focus · HN ↗
Just align it to do what the customer wants.
beej71 · · focus · HN ↗
augment_me · · focus · HN ↗
So the danger level is proportionate to how little money you have
tancky · · focus · HN ↗
Zealotux · · focus · HN ↗
disgruntledphd2 · · focus · HN ↗
We honestly don't know. Anthropic keep claiming this, and certainly have provided some evidence of this.
However, given that they've removed the real thinking output, and yet the Chinese models are still doing well, I'm sceptical of this belief.
cyanydeez · · focus · HN ↗
Aside from it being a completely hollow cry for attention from US labs, its also just what these models are designed around: taking in data and forming a way for them to operate on some level of intelligence and context. It's like a dictionary maker getting made that someone used their complicated words, and another dictionary maker heard the word and wrote it down then checked what the first dictionary maker wrote.
option · · focus · HN ↗
senectus1 · · focus · HN ↗
Killswitch Engineer
San Francisco, California, United States
300,000-500,000 per year
About the Role
Listen, we just need someone to stand by the servers all day and unplug them if this thing turns on us. You'll receive extensive training on "the code word" which we will shout if GPT goes off the deep end and starts overthrowing countries.
We expect you to:
• Be patient.
• Know how to unplug things. Bonus points if you can throw a bucket of water on the servers, too. Just in case.
• Be excited about OpenAI's approach to research
jswelker · · focus · HN ↗
reilly3000 · · focus · HN ↗
MiroslavPokorny · · focus · HN ↗
snorrah · · focus · HN ↗
... which is probably why they will go through with it. Losing a bunch of money seems to be one of the fundamentals of AI business here :(
reilly3000 · · focus · HN ↗
protocolture · · focus · HN ↗
1-6 · · focus · HN ↗
dlcarrier · · focus · HN ↗
paul7986 · · focus · HN ↗
Anyone else using Muse more and noticing similar stuff?
hosel · · focus · HN ↗
All I really would like is for some of you to CONSIDER THAT YOURE INCORRECT. Just imagine that people ringing the fire alarms are being sincere. Please entertain the position with an open mind.
kadoban · · focus · HN ↗
Did they just get a bunch of credulous "news" coverage out of it? Yes?
Will they just release this within the next weeks, at best? Yes?
Huh, funny that.
wbxp99 · · focus · HN ↗
antonvs · · focus · HN ↗
If all these claims were being made about some tech further along than LLMs currently are, they might be plausible. But LLMs on their own are not going to be “superintelligence” of the kind Altman is currently cynically spreading fear about. We know their limitations, and those limitations can’t simply be eliminated with more training or better harnesses.
> Just imagine that people ringing the fire alarms are being sincere.
The top three possibilities here are: they’re not being sincere, they’re just marketing; they’re being sincere, but they don’t understand the technology very well and are putting too much weight in what the first group are saying; they’re talking about a risk further in the future than OpenAI’s latest model.
No-one serious outside of OpenAI believes “this is the one”. At best, you’re conflating arguments being made on entirely different timelines.
MattPalmer1086 · · focus · HN ↗
From a corporate liability perspective, if your product is going to go off and hack loads of other companies, I would call it too dangerous (to the company) to release.
chrisjj · · focus · HN ↗
Dylan16807 · · focus · HN ↗
I don't believe they really care about danger, or that they'd actually fully withhold release on anything significant.
And when I say "significant", is 6.1 Astra even meaningfully different from 6 Astra in capability...?
literalAardvark · · focus · HN ↗
I don't get how this being an entirely legitimate issue is even hard to grasp.
Dylan16807 · · focus · HN ↗
literalAardvark · · focus · HN ↗
Letting people call those agents would expose them to a great deal of liability that just wasn't there in older models.
angoragoats · · focus · HN ↗
Glad you agree that OpenAI is responsible for this activity. Shouldn't we be charging the entire board of the company with, for example, violating the Computer Fraud and Abuse Act?
If this activity is so dangerous, and they developed the tools and continue to allow them to be used, it seems to trivially follow that they're knowingly violating federal law.
literalAardvark · · focus · HN ↗
Call your guy
angoragoats · · focus · HN ↗
blini-kot · · focus · HN ↗
chrisjj · · focus · HN ↗
angoragoats · · focus · HN ↗
I have considered that I’m incorrect, but I have literally zero evidence to indicate that the “very intelligent people” you speak of are correct, so I will maintain a position of non-belief of the claim until belief is warranted.
chasd00 · · focus · HN ↗
I’m not sure which logical fallacy this is but you’re not helping your case.
jiggawatts · · focus · HN ↗
GPT 4.5 was scrapped for similar reasons.
dash-44 · · focus · HN ↗
If you signed up for a 16 core AWS server and they randomly kept changing it down to 8 cores, you'd sue them. How long until the same applies to these AI companies.
The amount of meddling they do to the harness, system prompt, model, quantization etc makes these products sometimes unbearable; you never know what you're going to get. A few more iterations of Qwen 27B and hopefully we won't have to deal with any of this malarky any more.
jiggawatts · · focus · HN ↗
qalmakka · · focus · HN ↗
Clearly they've realised that people are increasingly fine with non-frontier models, including Chinese ones. They can't consolidate and just deliver a product to profit because they're constantly being caught up by the Chinese at prices they will never be able to match.
I suspect their main goal for the last year and a half or so was to pull a GM - let the share of the AI economy grow, then have the government bail them out when the debt becomes unsustainable. Clearly Trump would have never let the bubble explode before the midterms, right?
Well it's now clear that the US government has way too much debt right now for 2008-style bailouts, and the market doesnt have the money lying around to buy their overinflated stocks at an IPO. The only avenue realistically left to them is to cry wolf so much to get governments to ban all non-approved models, stop new developments and give them breathing room to consolidate to get their current crop of models profitable at higher prices.
Unfortunately for them I suspect Mr President has way too much dementia and way to little business acumen left to realise the sham going on, so I think they'll keep doing charades for a while until they'll suddenly stop
jaggs · · focus · HN ↗
moktonar · · focus · HN ↗
chrisjj · · focus · HN ↗
Let's try "you can only advance AI up to your highest level of competance."
In which case what we need is not a S-curve. It is an ∩-curve.
zerof1l · · focus · HN ↗
Aren’t all models doing this to some degree already? Ignoring some of the instructions, doing things beyond instructed, e.g., finding and fixing bug while doing something else. Especially Claude models. They seem to be in their own world with their own ideas about how things should be ran and done.
dash-44 · · focus · HN ↗
Opus and Sol are better for day to day dev work IMO in that they won't try to do too much.
ichorio · · focus · HN ↗
As in, I'd start Task A, mid-turn, I'd queue up "do B at the same time", and it'll accept it, but not do it. At the end, B just won't be done.
chrisjj · · focus · HN ↗
angoragoats · · focus · HN ↗
Or, if they’re not lying, then Sam Altman and the entire board of the company belong in a courtroom facing criminal for knowingly violating federal law.
Which is it, OpenAI?
chasd00 · · focus · HN ↗
conorcleary · · focus · HN ↗
rickydroll · · focus · HN ↗