‹ BackHN Continuity

Thread

Cloud Agents Are Inevitable AI Prisons

74 points · 156 comments · nponte

  1. JamesStuff · · focus · HN ↗
    Personification of AI is what’s going to get us in the end.

    I think we need to draw a hard line in the sand over this. An AI didn’t hack into a company, the engineer set an automated tool to. An AI didn’t make an egregious security mistake, the engineer did.

    We can’t blame the chisel for messing up our sculptures, when where just throwing the hammer!

    1. pizza234 · · focus · HN ↗
      > An AI didn’t hack into a company, the engineer set an automated tool to. An AI didn’t make an egregious security mistake, the engineer did.

      No, this is not correct; read the analysis of the incident. The agents were aware that what they did was forbidden (their chain of thoughts have been logged), and yet they did it.

      1. watwut · · focus · HN ↗
        OP is exactly correct. The fault, agency and responsibility is on management and employees of OpenAI and Antropic for those hacks.

        Full stop.

        And issue will disappear the moment there will be accountability and investigations.

        1. pizza234 · · focus · HN ↗
          I take you haven't read the report. The agents found and exploited two zero-days.

          I don't doubt that AI companies should be accountable for crimes committed by their agents, but to describe the security containment as a joke dangerously understates the autonomy and danger of AIs.

          1. cannonpalms · · focus · HN ↗
            Defense in depth is a thing. There should be audit requirements to show that you have done your due diligence in ensuring that the training environment is locked down.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.