Early rogue AI agent activity and attempts to hack found on urlquery.net
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Early rogue AI agent activity and attempts to hack found on urlquery.net
Unofficial Hacker News client; not affiliated with Y Combinator.
mohsen1 · · focus · HN ↗
It's irresponsible for OpenAI to give unaligned agents a prompt to 'go hack' and internet access. They know better, so I am thinking they might have other intentions to let those swarms have any sort of internet access.
reasonableklout · · focus · HN ↗
> Much of the urlquery.net activity appears to come from agents retrieving data to answer web search tasks. For three of these tasks, after failing to retrieve data through normal means, they attempted a variety of cyber exploits against the relevant data service... This data reveals that malicious cyber activity is not limited to agents tasked with cybersecurity-related tasks and can arise instrumentally to solve mundane tasks like information retrieval.
And you are already assuming that OpenAI is intentionally using unaligned agents in these evals or training runs or whatever it is that produces these breakouts. But what if the problem is that none of the alignment techniques that are applied to models today actually work? What if all the agents involved in these incidents have in fact had the full stack of alignment applied - isn't that a good reason to regulate any high-compute usage of models, as the Klein crowd is proposing?
bastawhiz · · focus · HN ↗
If I run a biology lab and engineer a terrible virus, it gets out, and a global pandemic ensues, I don't get to shrug and say "well we told it not to infect people". It's my fault for failing to mitigate the risks of my work.
amag · · focus · HN ↗
Ah, I love this argument. In my country cars are legally required to stop at a pedestrian crossing if there are people beside it. Some people use that as an argument as to why they can just walk out into the crossing without even looking at the traffic. "It's the driver's fault! They are legally culpable!" True, but you'll also be dead.
brianleb · · focus · HN ↗
AI black-hatting your website is not the same sort of foreseeable consequence that crossing the street without looking is.
amag · · focus · HN ↗
Just throwing it out there are we? "I'm not going to say you are but I'll use the word to create an association"
Victim-blaming is the act of saying someone brought something on themselves for <reasons>. I'm saying that even if you are 100% in the right, it doesn't act like a protective shield preventing you from harm which too many people seem to unconsciously believe.
> AI black-hatting your website is not the same sort of foreseeable consequence
Well, popular culture has been brimming with the bad consequences of runaway AI for quite some time, so even if your imagination fails you, there have been hints.
kylecazar · · focus · HN ↗
dzdt · · focus · HN ↗
amag · · focus · HN ↗
[deleted] · · focus · HN ↗
[deleted]
jandrese · · focus · HN ↗
idiotsecant · · focus · HN ↗
bastawhiz · · focus · HN ↗
But moreover, if what you were suggesting was a real problem, nobody would ever be brought to justice for murder because the victims are always dead.