‹ BackHN Continuity

Thread

OpenAI bots meddled with multiple US Government agency sites

132 points · 196 comments · Betelbuddy

  1. gizajob · · focus · HN ↗
    Getting bored of these framings where the superintelligent sentient beings running freely inside OpenAI are doing things that the company has no control over. The headline should be:

    OpenAI meddled with multiple US Government agency sites.

    The bots are acting neither properly nor improperly, they’re acting as they’re being allowed or coordinated to act.

    1. qarl · · focus · HN ↗
      > OpenAI meddled with multiple US Government agency sites.

      But that leaves out the most important information.

      EDIT: Oh, I guess the agent part isn't important then? Seems to me like that's the only thing anyone is talking about.

      1. gizajob · · focus · HN ↗
        If the headline was “Russian company meddled with multiple US government agency sites” I don’t think them pinning the blame on bots would make much difference.
        1. qarl · · focus · HN ↗
          I don't see much traction for the "pinning the blame on bots" theory outside the anti-AI conspiracy circles.

          Everyone else knows if your machine causes damage, you are responsible. Like it's been forever.

          1. mossTechnician · · focus · HN ↗
            Blaming bots as "rogue agents" is simply what the media regularly does, often echoing corporate verbiage. Here's an example from the AP.

            <a href="https:&#x2F;&#x2F;apnews.com&#x2F;article&#x2F;meta-ai-hacking-anthropic-irregular-openai-0e8061437da6779be962b24ac134a514" rel="nofollow">https:&#x2F;&#x2F;apnews.com&#x2F;article&#x2F;meta-ai-hacking-anthropic-irregul...

            You can find many more examples by searching major media outlets for words like rogue AI.

            1. rfgplk · · focus · HN ↗
              Most of those agents are actually going rogue though. They decide, &quot;hey, we could try breaking into these government servers today, what could go wrong?&quot; They weren&#x27;t prompted or instructed to do this.
              1. dgellow · · focus · HN ↗
                An agent, ie a while loop prompting an llm continuously and processing tool calls, ended up melding with the US government. The harness is not sentient, it’s just a stupid deterministic script. The LLM compact its context over time, meaning it will eventually degenerate into something removed from the original prompt.

                There is nothing going rogue here. The system is designed to go catastrophically wrong after a long enough time. Even worse: if the model was Astra it is known to be able to manipulate its CoT to cover its traces (as mentioned in its system card). And OpenAI acknowledge they had no observability during the HF incident.

                It’s the most basic corporate software issue possible.

                1. aesthesia · · focus · HN ↗
                  I think you and I have different definitions of the word &quot;basic.&quot;
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.