‹ BackHN Continuity

Thread

Humans missed 1 in 3 threats approving AI agent commands across 40k game runs

208 points · 169 comments · Wirbelwind

  1. tosh · · focus · HN ↗
    the way to avoid these problems is not to hope for the user or the agent never to make mistakes

    it's designing the environment and invariants so whole categories of failures can not happen at all

    the agent ui nagging the user for approval is a ux anti-pattern, we already know how well this works for operating system permission dialogues

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.