‹ BackHN Continuity

Thread

There are no "rogue" AI agents

396 points · 269 comments · zzzeek

  1. silverFork · · focus · HN ↗
    From what I understand, in one case, they had physically disconnected the sandbox from internet and asked it to do something and it had used connections through (import routines) that they had allowed, to pseudo escape the sandbox. Yes it wasn't obviously trying escape the sandbox but it escaped it because it doesn't understand the boundaries and neither do most humans other than the ones that provided the instructions that it had used. So it wasn't a rogue attempt but the fact that boundaries may be not be that easy to set despite what people think.
    1. zugi · · focus · HN ↗
      > physically disconnected the sandbox from internet ... used connections ... that they had allowed.

      That's not "physically disconnected the internet", that's "disabled some connections but enabled others."

      So the agent found and used the non-blocked connections.

      1. silverFork · · focus · HN ↗
        For the Ai code to execute, it needs the import functions... So that firewall between executing the code vs processing doesn't really work.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.