Who should be held accountable when an AI Agent (accidentally) acts maliciously?
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Who should be held accountable when an AI Agent (accidentally) acts maliciously?
Unofficial Hacker News client; not affiliated with Y Combinator.
CatDaaaady · · focus · HN ↗
Until we can agree whether AI is conscious, which we never will, AI and AI agents are just property working on behalf of humans.
I could see a future where AI companies/services indemnify consumers who use their agents but _not_ indemnify corporations that use their services.
trescenzi · · focus · HN ↗
qarl · · focus · HN ↗
I'm starting to wonder if the people arguing against anthropomorphization actually have any experience at all working with agents.
EDIT: It's a simple question. When you downvote me without answering, I must assume you don't have any answer and dislike what that implies.
trescenzi · · focus · HN ↗
Consensus is just populating the model with rationale for different new output.
qarl · · focus · HN ↗
I thought putting the question in my comment would be sufficient. I guess not. It was and still is:
> If I should not use anthropomorphic language, how do you suggest I handle the following situation
diegof79 · · focus · HN ↗
I've worked with agents, and I agree with you that often there isn't another way to express the interactions.
However, I also think the terms ML uses in general are a mimicry that misleads people who aren't informed. Ask anyone outside SWE what they think “training” means, and they'll usually picture something being taught.
I don’t think anybody can change that now, but it’s useful to point it out.
qarl · · focus · HN ↗
You mean how early-on people thought that planes flapped their wings while they flew?
These aren't problems. This is the way language works.
diegof79 · · focus · HN ↗
However, you can see an airplane flying. Still, you cannot see software processes at work, and that causes misunderstandings and misinformation, which is at the core of the changes that we are experiencing with AI.
This is a fragment of another article posted here on HN about an ongoing dispute between OpenAI and the New York Times:
“The defendants say this is a simple application of fair use: Their argument is that if you read a story and simply remember what was in it to expand your base of knowledge, that cannot be considered a copyright infringement”
However, if you replace “read” with “web scraping” and “expand your base of knowledge” with “storing the information,” the perspective changes too.
qarl · · focus · HN ↗
I think it's safe to trust the decisions of the courts. Judges aren't easily fooled by slippery language.
For example, in Bartz v. Anthropic, Judge Alsup ruled that training is fair use because training is transformative. In his words "spectacularly so".
nvme0n1p1 · · focus · HN ↗
I didn't downvote, but wow you're being aggressive, you have a lot to learn if you read the comments here with an open mind.
qarl · · focus · HN ↗
> If I should not use anthropomorphic language, how do you suggest I handle the following situation
And so strange that in a 24 hour period I got three comments all at the same time about being too aggressive and still not answering the question.
I'm sure it's just a coincidence.
nvme0n1p1 · · focus · HN ↗
The question is vague and doesn't seem related to the discussion. What do you mean "handle"? If you're having trouble handling it psychologically, see a therapist. If you're having trouble getting the output you want, look up guides on prompting. If you're anthropomorphizing the chatbot to the point you're worried about offending it... just don't worry? It's a computer program, don't overthink it, don't anthropomorphize it, just give it the input bytes you need to get the output bytes you want.
> I'm sure it's just a coincidence.
There isn't some grand conspiracy here. You just overindexed on the word "anthropomorphic" and didn't really understand what the discussion is about.
qarl · · focus · HN ↗
> I didn't downvote, but wow you're being aggressive
He said:
> I didn't downvote you, but your last paragraph is unnecessarily aggressive.
Both comments arrived within two minutes of each other on a thread with no other comments for 18 hours.
That's quite a coincidence.
Also - insults are against the rules here, friend. Naughty naughty. I hope they don't spank you.
nvme0n1p1 · · focus · HN ↗
This conspiracy goes all the way to the top. The illuminati assigned me personally to comment on your post. I've already said too much. If you never see me again, tell my wife I love her.
qarl · · focus · HN ↗
If it was a coincidence then you have nothing to worry about. I flag it because the overwhelmingly likely thing is that it wasn't.
And again - insults are against the rules here. Shame shame.