Who should be held accountable when an AI Agent (accidentally) acts maliciously?
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Who should be held accountable when an AI Agent (accidentally) acts maliciously?
Unofficial Hacker News client; not affiliated with Y Combinator.
throwitaway222 · · focus · HN ↗
So if an AI agent is asked to build a giant base for someone in MineCraft, and decided to build a swarm of additional agents, and one of those agents says "Time to destroy all humans" and autonomously hacks into the pentagon and fires the nukes - the company that developed the model is responsible. That being said - if the nukes deploy successfully, I have two questions:
1. If no one finds out, is anyone responsible?
2. Was any of this actually real?