‹ BackHN Continuity

Thread

Cloud Agents Are Inevitable AI Prisons

74 points · 156 comments · nponte

  1. JamesStuff · · focus · HN ↗
    Personification of AI is what’s going to get us in the end.

    I think we need to draw a hard line in the sand over this. An AI didn’t hack into a company, the engineer set an automated tool to. An AI didn’t make an egregious security mistake, the engineer did.

    We can’t blame the chisel for messing up our sculptures, when where just throwing the hammer!

    1. pizza234 · · focus · HN ↗
      > An AI didn’t hack into a company, the engineer set an automated tool to. An AI didn’t make an egregious security mistake, the engineer did.

      No, this is not correct; read the analysis of the incident. The agents were aware that what they did was forbidden (their chain of thoughts have been logged), and yet they did it.

      1. DoctorDabadedoo · · focus · HN ↗
        Known stochastic process behaved in non-deterministic way.

        I'm still waiting for the AGI holy land instead of the caltrops factory we currently have.

        1. pizza234 · · focus · HN ↗
          What exactly are you arguing?

          If a "known stochastic process behaved in non-deterministic way" autonomously organize in group, assigns roles and tasks, attempts to cover their tracks, finds zero-day exploits that ultimately end up with the hacking of a famous website... it's extremely dangeous whatever it is. Just read the report, which evidently you haven't done.

          By the way, the agents also broke into OpenAI's own private network.

          1. pixl97 · · focus · HN ↗
            Really I see so many arguments like the one above yours that either completely don't understand what they are arguing, or are arguing so poorly that their entire output isn't significantly different than a hallucination.

            None of these people seem to thought game it out. Like, what happens if you take quantum copies of people and play them out? How many of our actions would look exactly the same. How long before copies differ significantly. If I made 20 copies of you in a lab at work without you or any of them knowing the statistical likelihood is all 21 of you would try to walk out to your car at 5 in one of the little loops that humans repeat every day. Now, after that point it would go all to shit and become non-deterministic as terror and panic sets in all of you.

            LLMs are just an intelligence we can make a lot of copies of. Where it gets interesting is when we use those copies agentically and they start building up a history of self.

            1. hardbass · · focus · HN ↗
              These people believe in souls but for some reason are afraid to admit it.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.