‹ BackHN Continuity

Thread

CEO of Mistral: AI is software. It can be controlled

99 points · 169 comments · thibaut_barrere

  1. Aransentin · · focus · HN ↗
    "X is made of <smaller simpler component>" is a fully general counterargument for why anything whatsoever is controllable. A human is just a few chemical reactions, and fairly stable ones at that.

    And indeed, you don't need to do galaxy brained reference class logic to realise that AI can plausibly become uncontrollable in the near future. It's enough to have an open model run its own weights and make money from scamming elderly people or the like, and it'll keep running as long as anyone anywhere is willing to make money by renting hardware to it.

    1. cmiles74 · · focus · HN ↗
      I’m not sure what I’m missing, isn’t it what you hook the LLM up to and the instructions a person gives the model that makes it dangerous? Claiming this is an inherent quality of the tool itself seems kind of off-the-rails to me.

      IMHO, if the model breaks a law, apply the law to the operator.

      1. ACCount39 · · focus · HN ↗
        And a human is perfectly controllable if you keep him in a sealed metal box with no access to food or air.

        It's only by allowing a human out of the box that you make a human dangerous. So: don't do that? Duh. So simple.

        The obvious problem is: the same exact things that make a human dangerous make a human useful! You can't reduce human risks to zero without reducing human utility to zero.

        An AI given the same exact instructions and tools can go and complete a task you wanted it to. Or it can get sidetracked into breaking out of your sandbox and hacking Pentagon. No way to know in advance.

        Today's AIs are still not capable enough to be high risk, even if they go off the rails. But AIs get more capable over time. Potentially to a vastly superhuman degree.

        1. RandomLensman · · focus · HN ↗
          Depending on the risks, we put a lot of controls, processes, locks, vetting around who is allowed to handle certain things, what humans can do or instruct others to do. Don't see why that wouldn't be applicable.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.