‹ BackHN Continuity

Thread

Systems that no one will test

154 points · 84 comments · perone

  1. areoform · · focus · HN ↗
    The narrative around security and LLMs doesn't make sense to me.

    I think we've created a self-fulfilling prophecy. Everyone involved is acting with the best of intentions, but in avoiding what they fear, they've give shape and realized their fears. Much like a greek tragedy.

    An example of this is the story of Oedipus Rex, in the story Laius, the king, is told that he is "doomed to perish by the hand of his own son." (and wed his mother) And so to avoid this fate he decides to kill the infant. The person assigned to abandon him in the woods takes pity on the baby and gives the baby away. Thereby ensuring that Oedipus knows neither his mother or his father (and arguably giving him a reason to kill his father).

    The child grows up and hears the same prophecy again and the child tries to avoid the prophecy as well, as he loves his adoptive parents. So he leaves them and travels to Laius' kingdom, where he runs into Laius. Neither recognizes the other. As Laius is the type of man to kill an infant, they end up in an argument, whereupon Oedipus kills him.

    I think the ancients were on to something, because if Laius had reacted to the prophecy with courage, he would have been saved. I would like to argue that if he had faced his fear and raised Oedipus with love, then the necessary preconditions for the prophecy to come true wouldn't have taken root. But that's not what happens.

    By being driven by his neuroses and in acting with cruelty out of fear, Laius makes the prophecy real.

    To quote Heraclitus, ethos is fate. Or, character is fate.

    I think a lot of people in this AI research sub-culture would be served well by reading these classics, because they are making their self-prophesied doom come true.

    They have been convinced for years (GPT-2 was released in Feb 2019) that AI is dangerous. A tremendous threat. An apocalyptic threat.

    One dimension of this fear has been the idea that a super smart AI will take over our digital infrastructure and be responsible for the digital apocalypse. That would be terrible!

    So what do they do?

    They try to make a counter to their fears by teaching models how to exploit vulnerabilities.

    How dangerous is such an entity? Very!

    Convinced of this danger, they start testing their models as if they were weapons with offensive capability. And then they create models that can be used as weapons.

    And because they don't want to release a dangerous weapon out into the world (oh no!), they restrict access to their AI, thereby depriving everyone of tools they can use to improve their security...

    Ethos anthropoi daimon.

    1. myrmidon · · focus · HN ↗
      I don't disagree with your main point.

      But this attitude towards dangerousness of AI I simply don't understand.

      In my view, AI is obviously risky and dangerous, because it is the only thing on the planet that can think/reason at a human level (or higher) apart from us (and that capability is what made us the uncontested apex species on the planet).

      AI is unconstrained by hard biological limits; keeping up with its capabilities will be impossible for baseline humans.

      You can argue all day about particulars, like whether selfreplicable robotic shells are required to reach/exceed "existential risk" level to our species (or if artificial minds are enough of a potential threat by themselves).

      But the general "risk" is simply AI becoming able to act in its own interests to our detriment, and I don't understand how anyone can dismiss this right now-- help me understand.

      Of course there are plenty of imaginable scenarios where coexistence is peaceful and mutually (?) beneficial, but that does not answer those concerns by itself at all.

      1. vladms · · focus · HN ↗
        I find the discussion about AI starting from the (mostly) for profit companies involved in AI very dubious. There are a many fields (bio-tech, nuclear, chemical) which could provide an "existential risks" to the human species, but most of the companies involved argue they control it, while other people argue they don't do it well enough.

        I do think AI is posing a threat, but not "by itself" but mostly what it can do in the hands of misguided (or evil) people. The same as with previous risks.

        And if our digital infrastructure can be brought down by something (AI/human/government) it is just badly done, being afraid of "AI" will not improve the situation. Someone should come with a proposal about what do we do now, but I would prefer more root cause solutions (how to make resilient systems) rather than "fear of AI". But building is hard and fear catches media attention, so...

        1. myrmidon · · focus · HN ↗
          > There are a many fields (bio-tech, nuclear, chemical) which could provide an "existential risks"

          I honestly don't think chemical and current biological weapons even register on the extinction-risk scale.

          All out nuclear war I would put at the very low end of the scale (likely to end our current civilization, very unlikely to wipe out our species).

          The big risk with AI I'm anticipating is not it being abused for desinformation campaigns or cyberattacks; the big risk in my view is AI actor(s) consolidating manufacturing control, obtaining some kind of self-replicability and physical agency and then just pursuing self-interests (within the next century). There is a lot of imaginable scenarios where that simply ends us, as a species, completely.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.