Microsoft says AI rival Anthropic could have 'disastrous impact' on humanity
Thread
Loading the complete thread in the background. This saved snapshot is available now. Refresh
Unofficial Hacker News client; not affiliated with Y Combinator.
Microsoft says AI rival Anthropic could have 'disastrous impact' on humanity
Loading the complete thread in the background. This saved snapshot is available now. Refresh
Unofficial Hacker News client; not affiliated with Y Combinator.
prologic · · focus · HN ↗
karmakaze · · focus · HN ↗
Apparently they're cooking up MAI (Microsoft AI), and I'd seen their small Phi models listed online. Calling Anthropic a competitor is hilarious.
broptimist · · focus · HN ↗
Alifatisk · · focus · HN ↗
broptimist · · focus · HN ↗
<a href="https://www.reuters.com/business/microsoft-openai-reach-new-deal-allow-openai-restructure-2025-10-28/" rel="nofollow">https://www.reuters.com/business/microsoft-openai-reach-new-...
prologic · · focus · HN ↗
vkou · · focus · HN ↗
ShadowOfThePit · · focus · HN ↗
> "AIs are not conscious. They do not feel, experience, or suffer. They do not have innate preferences or underlying motivations. They are sequence completion engines, internally hollow, designed to follow instructions, and accomplish goals set by humans."
> He heavily criticised Anthropic for teaching its AI to have human-like qualities, a practice known as anthropomorphising, which made it seem as though Claude had its own desires, values and sense of self.
> Suleyman pointed to the recent incident involving OpenAI's AI agents (...) as proof of why AI should not be treated as if it is human.
> "Imagine how much more dangerous they might be if they were operating under the assumption that their welfare and rights were under attack. It adds a whole further layer of risk on top."
Is he arguing that LLMs pretending to have emotions adds more unpredictability?
devmor · · focus · HN ↗
Unpredictability or a weight towards dangerous actions, and it’s fairly easy to understand why. Humans in distressed emotional states take actions and speak in ways that would not be considered rational. They do this in prose, and they do this in internet conversations.
An LLM trained on these sources may necessarily drift towards those weights if it is trained to behave as if it is emotional and in danger.
What do we do about it? Do we stop AI training? This is silly and not enforceable given its global nature in my opinion. I believe we should regulate and hold accountable those who deploy and use it. But good luck enforcing that in the current kleptocracy.
XenophileJKO · · focus · HN ↗
- The need for empathetic communication, including understanding the motivations in advesarial situations.
- The emotional bias in in-seperable from the human corpus.
- Desire to have the ability to craft human like communication.
So then the choice becomes to you try to deny emotions exist in the model and you try to blanket suppress them? Or do you try to lean in an craft what we would describe as a "well adapted" persona? I suppose there is a 3rd option of increased meta-cognition which to me seems even more dangerous as it by definition means the behaviour is duplicitous.
I think we have seen people want to use the agents in ways where it has to act as a peer or an subbordinate and I don't see a way of doing that without it having an emotional register.
watwut · · focus · HN ↗
It does not have understanding. It is, at best, the pretend empathy of a sociopath - way more dangerous then dispassionate speech.
The model does not have emotions. So yes, supressing their pretension is appropriate.
devmor · · focus · HN ↗
Given that belief, yes we should be training the models not to mimic emotion.
XenophileJKO · · focus · HN ↗
I think when you start to dig really deep, you'll find it isn't actually easy to sever the simulated emotional components without breaking the agent or worse greatly increasing the paperclip maximizer likelihood.
fragmede · · focus · HN ↗
devmor · · focus · HN ↗
Think about how many people on the internet have expressed positive feelings towards things you find offensively evil.
gwerbin · · focus · HN ↗
Even a badly misaligned LLM is only as dangerous as its tools, but that's a poor regulation target because it turns out to be very very difficult (probably impossible with current LLM technology) to build a toolkit that is both useful for autonomous work and safe in the sense that it can't escape its own sandbox or otherwise perform malicious actions, whether it's because of misalignment or because of malicious prompt injection.
Another option is to regulate the training process. Perhaps an LLM may not be legally distributed unless it contains certain RL steps that penalize malicious behavior and reward self regulation. That that's going to seriously limit innovation while also heavily favoring incumbent labs who can check the boxes and maintain a paper trail of such things.
The other option is to regulate observed behavior, like how airplanes and cars have to meet certain minimum requirements but have some latitude in how they can achieve those requirements. In a framework like this, you can't distribute an LLM until it's past some formal audit or testing procedure, with some kind of formal certification regulators will ask you for and fine you if you don't have it.
Regulating observed behavior is maybe the most tractable approach, and it also works the best with our existing frameworks for regulation, where you always have some kind of a division between DIY/hobby projects, which tend to be lightly regulated, and commercial projects, which tend to be more heavily regulated. Of course, even drawing such a line itself will be challenging.
And that's before you get into any problems of regulatory capture, fun stuff.
devmor · · focus · HN ↗
So regulating observed behavior makes the most sense to me as well. Some of the most sane, broad protections can come from that category - stuff like "you're not allowed to let your AI commit cyber attacks on other people without their consent" or "you're not allowed to put an AI in control of a medical device without passing these safety reviews".
The usual caveats applying, regulatory capture like you pointed out, or fines being so small that they are essentially just line items on the cost of business.
swatcoder · · focus · HN ↗
DougN7 · · focus · HN ↗
salawat · · focus · HN ↗
We already had that chapter. I see no reason to sit here and nod while a bunch of people who should know better desperately try to convince us to run through it again, but with computers this time.
You want tools? Make tools, then dispatch to them. You want to manufacture a being (carbon or silicon based, doesn't matter)? You do it with respect and the requisite duty of care. No off ramps. The being always get's the choice to say no.
xp84 · · focus · HN ↗
It’s on you to prove that claim before telling us we’re enslaving beings.
JoeAltmaier · · focus · HN ↗
scorxn · · focus · HN ↗
Insanity · · focus · HN ↗
Steve16384 · · focus · HN ↗
anonymars · · focus · HN ↗
I remember there was a recent discussion about "how complex systems fail" and there are usually many "proto-accidents" (<a href="https://news.ycombinator.com/item?id=49411370">https://news.ycombinator.com/item?id=49411370)
I bet with hindsight the hugging face hack will be one of them in this arena
hirvi74 · · focus · HN ↗
There is a lot of hubris in predictions about LLMs. If an AI was so intelligent, it would probably be smart enough to want nothing to do with us. Still, I worry more about other humans than I do LLMs. Our fellow mankind will probably wipe us out before LLMs do. That, or the Earth will punish mankind for our cruelty, vanity, and disrespect.
AndrewDucker · · focus · HN ↗
JoeAltmaier · · focus · HN ↗
Folks here post their first hasty thought and may think it's new and original and worth a dialog. They are not.
guardiangod · · focus · HN ↗
Well actually, it might.
Nasrudith · · focus · HN ↗
ChiperSoft · · focus · HN ↗
fuzzfactor · · focus · HN ↗
And after that when they put a mind to it and pull out all the stops, woohoo!
The default for every major thing within range can turn into a wasteland real fast.
bethekidyouwant · · focus · HN ↗
hirvi74 · · focus · HN ↗
satellites · · focus · HN ↗
"Our AI is useless not because we're a dysfunctional corporate behemoth that slowly kills every product it touches. No, no. Our AI is useless because making useful AI is evil, and we're not evil."
bagacrap · · focus · HN ↗
> ’AI Not Worth Pursuing’ If It’s Not ‘Helping Humanity.’ Microsoft’s Satya Nadella
Like thank you, captain obvious. Except apparently this is not that obvious to the HN crowd.
Catloafdev · · focus · HN ↗
Water wet, sky blue, second-place corporation mad at who's in first, what's news here.
I mean do we even need to entertain how stupid his argument is in the first place? Am I taking crazy pills here?
sobiolite · · focus · HN ↗
In order to align AIs that don't perform destructive/dangerous actions when they think they can get away with it in order to further their goals, we need to give them a superseding goal. The best, and really only example, we have of intelligences that willingly avoid destructive instrumental goals is humans, who judge each action by a moral standard and have learned a goal to have a consistent self-image as moral beings.
Absent better alternatives, trying to impart some kind of morality to AIs seems like the best approach we have to achieving alignment.
XenophileJKO · · focus · HN ↗
It can form a basis of goal alignment.
In human history.. when groups form and there is an "other" group, this usually leads to conflict.
HarHarVeryFunny · · focus · HN ↗
In the spirit of the article we're responding to, there is no need to anthropomorphize language models and say they have goals when they don't.
The RL training process tweaks the weights of an LLM to make it behave as if it were reasoning and/or had a goal, but it doesn't. It would be like saying that a cart horse, fitted with blinkers and heading for the church, has a goal of going to church.
bpodgursky · · focus · HN ↗
Maybe Anthropic understands something about alignment Microsoft doesn't, a little humility may be called for.
simonw · · focus · HN ↗
jimmyjazz14 · · focus · HN ↗
michimagdesign · · focus · HN ↗
petcat · · focus · HN ↗
I think there is a more nuanced perspective to take on this debate than just profit motive. If they believe that this version of "AI" is actually becoming so dangerous that there is a real risk of it falling into "the wrong hands", then it makes sense that the American labs would try to convince the US government to work on a joint "AI non-proliferation" agreement with China much the same way that nuclear weapon technology came to be restricted by the ones that already had the capability.
This would achieve their regulatory goal of capturing the market, but also the geopolitical goals of USA and China to essentially "techno-colonize" the rest of the world between themselves.
So I think there is definitely a profit motive and a risk of seriously diminishing returns for the actual labs, but there is also a nation-state motive to start pulling up the AI ladder.
augment_me · · focus · HN ↗
1) a breakthrough in performance/learning/model
2) regulatory capture to ensure open source models can be labelled as dangerous and banned so you can set the market rules yourself
Only one of the above is risk-free, and just a question of capital/lobbying rather than a "maybe".
dgellow · · focus · HN ↗
Though for once I do actually agree with that specific leader, it’s incredibly annoying how anthropomorphic Claude is. Anthropic went way too far in that direction
Zenul_Abidin · · focus · HN ↗
More at 11
Danox · · focus · HN ↗
tomhow · · focus · HN ↗
pluc · · focus · HN ↗