‹ BackHN Continuity

Thread

OpenAI halts training of latest models as reports mount of AI agents going rogue

59 points · 118 comments · smb06

  1. hbarka · · focus · HN ↗
    ‘There are no “rogue” AI agents’

    <a href="https:&#x2F;&#x2F;eoinhiggins.substack.com&#x2F;p&#x2F;there-are-no-rogue-ai-agents" rel="nofollow">https:&#x2F;&#x2F;eoinhiggins.substack.com&#x2F;p&#x2F;there-are-no-rogue-ai-age...

    1. digitaltrees · · focus · HN ↗
      The article you link to is wrong, it states &quot;AI cannot think for itself, nor can it take independent actions.&quot; This is flawed reasoning, AI doesn&#x27;t need to &quot;think&quot; in the way humans do to have autonomy. Go to codex or claude code or any harness right now, type a prompt and see if it executes a bash command or web search or file edit that you didn&#x27;t tell it to, that is an autonomous plan and execution. If anything it&#x27;s even more dangerous that thet can call drop db or kill pid without a user giving instructions.
      1. jpnc · · focus · HN ↗
        &gt;AI autonomy

        &gt;&#x27;type a prompt&#x27;

        Which is it?

        1. digitaltrees · · focus · HN ↗
          You’re confusing time series. If my prompt is “add twilio sms sending to the app” and the AI logs into railway, reads my credentials, gives them to a subagent that in turn posts it to a message board resulting in my credentials being publicly available on the internet, those intermediate actions don’t reasonably follow from my instruction. No reasonable person would expect that series of events and the AI is fully capable of planning them, deciding to execute them, and deciding to report or hide them from me. If that isn’t autonomous then what is.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.