‹ BackHN Continuity

Thread

Nvidia wants to put a watchdog chip next to every AI agent

230 points · 299 comments · jonbaer

  1. cedws · · focus · HN ↗
    A new chip solves nothing. Nobody wants to hear this but there is no solution for the security risks posed by agents today. You can put it in a sandbox, it doesn't make a difference, for it to be useful it inherently needs wide, unattended access. Put a human in the loop and you just end up bottlenecking it and throwing away any purported productivity gains. Auto mode doesn't matter either, it's trivial to trick and for the agent to break out.
    1. nicce · · focus · HN ↗
      > Put a human in the loop and you just end up bottlenecking it and throwing away any purported productivity gains. Auto mode doesn't matter either, it's trivial to trick and for the agent to break out.

      Productivity gains are still enormous compared to what we used to do before agents. But, I know that people don't want to stop there.

      1. paimapi · · focus · HN ↗
        right, the solution here is not a hyper-capitalist race-to-the-bottom-of-devaluing-labor. it's recognizing discretion and diligence are things still required for work to be of a certain quality
      2. egeozcan · · focus · HN ↗
        Humans can also be tricked by the agents.

        Humans can be tricked by humans too but humans care about their reputation in their communities, and at least fear from punishment.

        1. gus_massa · · focus · HN ↗
          Computer says no has been a problem for decades. The human can blame the computer for the errors following it, but must assume the consecuences if they override the decision.
          1. [deleted] · · focus · HN ↗

            [deleted]

          2. wavewrangler · · focus · HN ↗
            "I don't want the details"
          3. intended · · focus · HN ↗
            Individual responsibility is meaningless when talking about a system and economy level change.

            Unless something is in the structure that makes individual choice and responsibility a meaningful source of friction and reduced velocity, it has no real impact on how AI is being used.

      3. autoexec · · focus · HN ↗
        > Productivity gains are still enormous

        Depends on who you ask I guess

        <a href="https:&#x2F;&#x2F;www.theregister.com&#x2F;software&#x2F;2026&#x2F;01&#x2F;15&#x2F;ai-is-everywhere-but-nowhere-in-recent-productivity-data&#x2F;4845104" rel="nofollow">https:&#x2F;&#x2F;www.theregister.com&#x2F;software&#x2F;2026&#x2F;01&#x2F;15&#x2F;ai-is-everyw...

        <a href="https:&#x2F;&#x2F;www.zdnet.com&#x2F;article&#x2F;workslop-can-kill-your-productivity-heres-how-to-turn-ai-into-a-competitive-advantage&#x2F;" rel="nofollow">https:&#x2F;&#x2F;www.zdnet.com&#x2F;article&#x2F;workslop-can-kill-your-product...

        <a href="https:&#x2F;&#x2F;fortune.com&#x2F;2026&#x2F;08&#x2F;22&#x2F;executives-ai-productivity-layoffs-study&#x2F;" rel="nofollow">https:&#x2F;&#x2F;fortune.com&#x2F;2026&#x2F;08&#x2F;22&#x2F;executives-ai-productivity-la...

        1. tniemi · · focus · HN ↗
          It&#x27;s probably just productivity paradox v2. <a href="https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Productivity_paradox" rel="nofollow">https:&#x2F;&#x2F;en.wikipedia.org&#x2F;wiki&#x2F;Productivity_paradox

          It takes time for decades old ways to change.

      4. realusername · · focus · HN ↗
        &gt; Productivity gains are still enormous compared to what we used to do before agents.

        My own productivity yes but if I step back and look at a company scale, the productivity gains has been negative for our company as a data point.

        We are now shipping less and with a lower quality.

        1. 8n4vidtmkvmk · · focus · HN ↗
          How are you shipping less?

          I believe the lower quality but is quality so bad you are afraid to ship it now?

          1. krageon · · focus · HN ↗
            You cannot ship things that don&#x27;t work unless you work for Microsoft or Oracle or I guess IBM
            1. hdjrudni · · focus · HN ↗
              I guess, but why are you even submitting&#x2F;merging code that doesn&#x27;t work? Like how does it even get that far that you&#x27;re blocking other people from getting stuff done?

              I guess I won&#x27;t deny that we&#x27;ve had more build breakages than usual over the past couple months but they get resolved super quick. We don&#x27;t hesitate to roll broken code back if something manages to slip in.

          2. realusername · · focus · HN ↗
            AI also had a negative impact on the CI and on time spent to review code &amp; documents so because of that, we&#x27;re also shipping less
        2. intended · · focus · HN ↗
          V&#x2F;G, the ration of verification and generation capacity is borked in AI using firms now.

          It’s not an issue of only more generation, it’s an issue of how much generation outstrips capacity to verify generated content.

          Unlike spam, you can’t filter out and bin the stuff a colleague is sending you.

          So individual productivity is up, while the costs of checking and processing generated content shifted to the rest of the org.

      5. dgellow · · focus · HN ↗
        So much productivity gain, and yet still zero proof of positive contribution to companies ROI. Unless you’re yourself reselling AI of course
      6. SkyBelow · · focus · HN ↗
        Are they? How many people claiming productivity gains are actually being a human in a loop and reviewing and understanding every code change and every line of code ran?

        2 years ago I saw it, back before agents were really a thing. But I&#x27;m not sure that was an enormous productivity improvement, especially compared to agents today. As for today, everyone I talk to is some level of blindly trusting what the agent is doing or not using agents. I haven&#x27;t met people in the middle ground and suspect that they are rare enough we don&#x27;t really know what their productivity gains are.

        Human in the loop has become a convenient security-theater-washing for agentic AI.

        Outside of coding, I think the issue is even worse because humans will defer so much judgment to AI that the same would apply. Look at how much trouble we&#x27;ve had before modern AI where humans blindly trusted the computer&#x27;s output rather than make their own judgment, even if their job was to be providing a safeguard against the computer&#x27;s judgment.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.