‹ BackHN Continuity

Thread

Humans missed 1 in 3 threats approving AI agent commands across 40k game runs

208 points · 169 comments · Wirbelwind

  1. continuational · · focus · HN ↗
    It's kinda funny there is still software coming out whose security model is "constantly ask the user for permission, and hope they never make a mistake".

    It's been tried so many times before, and it never worked.

    1. ApolloFortyNine · · focus · HN ↗
      Air Traffic Control is still primarily voice based, and simply up to the user on both sides to not make a mistake.

      Just bringing it up because you're right, in software that's considered a bad pattern (rightfully so).

      1. overfeed · · focus · HN ↗
        > Air Traffic Control is still primarily voice based, and simply up to the user on both sides to not make a mistake

        The "user[s] on both sides" of ATC conversions have passed through the filters of rigorous training and certification. They also happen to communicate in a DSL designed to minimize misunderstandings, the DSL just happens to be based on English.

        1. ApolloFortyNine · · focus · HN ↗
          I don't think OPs argument here was simply that the users aren't qualified enough to approve llm output.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.