OpenAI Discloses Six New Incidents of ‘Concerning’ A.I. Behavior
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
OpenAI Discloses Six New Incidents of ‘Concerning’ A.I. Behavior
Unofficial Hacker News client; not affiliated with Y Combinator.
teagee · · focus · HN ↗
Would nytimes cover a self driving car company disclose concerning ‘behavior’ of their cars the same way?
For anyone who has had to remind a coding agent to not leave comments over and over again, not following instructions seems more feature than bug
bitexploder · · focus · HN ↗
[deleted] · · focus · HN ↗
[deleted]
EA-3167 · · focus · HN ↗
bigglebear · · focus · HN ↗
"Our unreleased model attempted to create a bioweapon", but "trust me bro, we didn't tell it to do that. We didn't train the model on a dataset that specializes in creating and glorifying bioweapons. We'd never stand to gain from misleading people about model capabilities in any way shape or form." - Anthropic are renowned for doing exactly this, for starters.
So this ends up resulting in more safety theater. You can't have anything fruitful come of this without transparency. Stop trying to protect your moat if you truly care about safety and actionable outcomes, and provide real transparency, otherwise this is as good as saying nothing at all.
I'm not even saying they're intentionally trying to do this by the way, but this is not sufficient if the goal is balanced incentives and accountability.
swat535 · · focus · HN ↗
Basically, the billionaire elites and their employees are panicking because they fear their funds are at risk when the bubble bursts.