‹ BackHN Continuity

Thread

There are no "rogue" AI agents

396 points · 269 comments · zzzeek

  1. silverFork · · focus · HN ↗
    From what I understand, in one case, they had physically disconnected the sandbox from internet and asked it to do something and it had used connections through (import routines) that they had allowed, to pseudo escape the sandbox. Yes it wasn't obviously trying escape the sandbox but it escaped it because it doesn't understand the boundaries and neither do most humans other than the ones that provided the instructions that it had used. So it wasn't a rogue attempt but the fact that boundaries may be not be that easy to set despite what people think.
    1. verdverm · · focus · HN ↗
      You mean the proxy to package registries from one of the early incidents?

      I have not heard about any instances where physical disconnect has happened, would appreciate any links to update my priors

      other non hacking cases of negligence include suicide and school shootings, which I have heard they were aware of and monitoring, but did not contact authorities

      1. silverFork · · focus · HN ↗
        There are references about escaping offline sandbox.. I dont know about shootings!!

        <a href="https:&#x2F;&#x2F;www.primeintellect.ai&#x2F;blog&#x2F;universal-offline-sandbox-escape" rel="nofollow">https:&#x2F;&#x2F;www.primeintellect.ai&#x2F;blog&#x2F;universal-offline-sandbox...

        1. hn8726 · · focus · HN ↗
          If you read the article it states clearly that the sandbox wasn&#x27;t offline though? There were API calls to certain endpoint(s) allowed, and the model simply used that endpoint&#x27;s feature to query data from the internet
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.