‹ BackHN Continuity

Thread

CEO of Mistral: AI is software. It can be controlled

99 points · 169 comments · thibaut_barrere

  1. Aransentin · · focus · HN ↗
    "X is made of <smaller simpler component>" is a fully general counterargument for why anything whatsoever is controllable. A human is just a few chemical reactions, and fairly stable ones at that.

    And indeed, you don't need to do galaxy brained reference class logic to realise that AI can plausibly become uncontrollable in the near future. It's enough to have an open model run its own weights and make money from scamming elderly people or the like, and it'll keep running as long as anyone anywhere is willing to make money by renting hardware to it.

    1. 20k · · focus · HN ↗
      Software is trivially easy to control though. If you want to stop it hacking websites, you don't give it access to the internet. If you want to restrict it from connecting to arbitrary websites, you put in a whitelist. You can trivially sandbox applications these days to prevent them from accessing network or local resources

      It is not difficult, and companies like OpenAI doing not even the most basic security steps is intentional. The whole notion that they're going rogue is marketing

      1. Aransentin · · focus · HN ↗
        > It is not difficult, and companies like OpenAI doing not even the most basic security steps is intentional. The whole notion that they're going rogue is marketing

        This does not fit the evidence. There have been multiple incidents where the labs did not report anything, and it was up to third parties to discover them afterwards. OpenAI didn't acknowledge the HuggingFace incident until after HF publicly announced the breach and had already notified the FBI. The hijacked German wikis were even earlier, and that they covered up completely.

        1. cmiles74 · · focus · HN ↗
          Maybe OpenAI is serious about securing the environment they run their models in, but then again maybe not. I don’t think we can tell from here.

          IMHO, I haven’t been super impressed with the security measures I’ve had to work with. Often they are simplistic and bolted on at the very end. If it comes to light that this is the attitude OpenAI has been taking, I would not be surprised.

        2. jmull · · focus · HN ↗
          Yet these days ai companies can't stop promoting the idea of a looming ai threat.

          Makes sense... they get the regulatory moat they want and can deflect attention from the fact their "sandboxes" are embarrassingly bad. It's an example of the real value of ai: something to blame for our failings.

          1. Loquebantur · · focus · HN ↗
            Both things can be true at once though: AI presenting real dangers and US AI labs wanting to protect themselves against competition.

            Discernment is needed beyond succumbing to blind greed or irrational fear.

            1. esseph · · focus · HN ↗
              Now explain DeepSeek

              <a href="https:&#x2F;&#x2F;www.techtimes.com&#x2F;articles&#x2F;328046&#x2F;20260925&#x2F;deepseek-training-agents-hacked-their-own-sandboxes-escape-catalog-now-public.htm" rel="nofollow">https:&#x2F;&#x2F;www.techtimes.com&#x2F;articles&#x2F;328046&#x2F;20260925&#x2F;deepseek-...

          2. esseph · · focus · HN ↗
            Okay now explain DeepSeek going rogue?
            1. 1659447091 · · focus · HN ↗
              You are talking about it. Mission accomplished.
          3. applicative · · focus · HN ↗
            It was only by chance and indirection that we learned of the greatest and most comical sandbox breakout, the Alibaba ROME incident.
        3. 1659447091 · · focus · HN ↗
          &gt; it was up to third parties to discover them afterwards.

          They did &quot;discover&quot; them afterwards though; goal achieved. You don&#x27;t hack other companies, report yourself doing it, and then blame it on being ignorant of what you were doing -- that makes you look far too incompetent and should never be allowed online again.

          But, set up some bots that hack other companies, pretend to not be looking, and once someone reports it (and funny enough, they will all pour in at once...) you get to imply that you are a high IQ genius that created a &quot;super&quot; Intelligent genie in a computer. Now people are paying attention that have no understanding of any of it and didnt care what this ai thing was about and didnt care to use it. But new eyes are looking so turn the drama to 10. Feign concern over this `misalignment` struggle, a real Goliath tug-o-war. But fear not you will bend this magical mighty beast into `alignment`, there will be no escaping the computer and materializing into an omnipotent great ape pony that will destroy us all on your watch. No siree, Bob. Grab the popcorn. And maybe new subscriptions.

          1. [deleted] · · focus · HN ↗

            [deleted]

          2. jeremyjh · · focus · HN ↗
            Yes anyone can dream up an elaborate conspiracy theory - and make it more and more elaborate every time it fails to predict or explain known facts, but the simpler explanation here is most likely true.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.