‹ BackHN Continuity

Thread

Nvidia wants to put a watchdog chip next to every AI agent

230 points · 299 comments · jonbaer

  1. wavewrangler · · focus · HN ↗
    Did they try just properly sandboxing them first? Or are they still learning how to configure a firewall over there?

    The problem isn't even the AI, the problem is the people in charge of the AI. This is a fabricated crisis

    1. pyronite · · focus · HN ↗
      > The problem isn't even the AI, the problem is the people in charge of the AI. This is a fabricated crisis

      This is a very confident statement in the face of a purported non-0% chance of human extinction.

      For what reasons do you disagree with the dangers of an intelligence explosion, e.g. Geoffrey Hinton and other experts in the field? <a href="https:&#x2F;&#x2F;www.theguardian.com&#x2F;technology&#x2F;2026&#x2F;sep&#x2F;28&#x2F;ai-godfathers-warn-of-runaway-intelligence-explosion" rel="nofollow">https:&#x2F;&#x2F;www.theguardian.com&#x2F;technology&#x2F;2026&#x2F;sep&#x2F;28&#x2F;ai-godfat...

      I&#x27;m curious why you and others seem to write off the possibility so strongly. I would love to feel more confident.

      1. voidhorse · · focus · HN ↗
        There&#x27;s a difference between the current material risks (which OP correctly identifies reduce down to basic human incompetence) and the long term hypothetical risks (which is what Hinton is concerned about).

        There are clear procedures for dealing with the immediate risk that have been known to the software industry for a long time. Don&#x27;t let the companies use hypothetical risks as a smokescreen to hide their negligence.

        1. reasonableklout · · focus · HN ↗
          Nobody is saying we should not hold the labs liable for damages caused by their negligence.

          At the same time, the technology is advancing in capabilities exponentially, and is beginning to exhibit long-predicted failure modes of RL that are nevertheless quite different than “insecure sandbox” or other that the software industry is used to.

          The current crisis which OP claims is “fabricated” comes from the fact that the technology is advancing faster than anyone anticipated, the Hugging Face incident provides a clear example everyone can point to, and the labs have realized they cannot self-regulate because of a collective action problem.

          There are a lot of levels of catastrophic damage that can happen between now and “long term hypothetical risks” like human extinction. When will it be worth regulation for you?

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.