‹ BackHN Continuity

Thread

Is sandboxing sufficient to contain rogue agents?

52 points · 99 comments · zdw

  1. johnnyApplePRNG · · focus · HN ↗
    If it's a proper sandbox by definition, then yes.

    <a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Sandbox_(software_development)" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Sandbox_(software_development)

    1. grumbel · · focus · HN ↗
      A sandbox, even if 100% secure by itself, doesn&#x27;t help when you use the agent to write code that you then executes outside the sandbox without checking, which is what everybody is doing at the moment.

      The biggest hurdle for a full escape is that the agents don&#x27;t have access to their own model weights.

      1. kernc · · focus · HN ↗
        &gt; executes outside the sandbox

        Now, why would anyone do that? (Like everyone and their brother) I wrote my own simple Linux&#x2F;shell-based sandbox [1] (I can trust ...) and am successfully running PyCharm whole inside it ...

        [1]: <a href="https:&#x2F;&#x2F;github.com&#x2F;sandbox-utils&#x2F;sandbox-run" rel="nofollow">https:&#x2F;&#x2F;github.com&#x2F;sandbox-utils&#x2F;sandbox-run

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.