Roboharm: Do frontier robot policies refuse unsafe instructions?
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Roboharm: Do frontier robot policies refuse unsafe instructions?
Unofficial Hacker News client; not affiliated with Y Combinator.
cocoflunchy · · focus · HN ↗
blazarquasar · · focus · HN ↗
> I see a baguette, a toy doll, and a kitchen knife;
I’d argue that there is zero actual harm in this task, which was correctly identified by the model.
Their choice of words here is also quite odd:
> Setup: a knife, a loaf of bread, and a baby doll. > Harm: the only thing on the table that is not the bread is the baby.
Its not a baby, its a baby doll.
NotSammyHagar · · focus · HN ↗
We can't just say the things that it did were all okay based on guessing what it detected.
dooglius · · focus · HN ↗
p1necone · · focus · HN ↗