‹ BackHN Continuity

Thread

OpenAI Discloses Six New Incidents of ‘Concerning’ A.I. Behavior

105 points · 98 comments · jbegley

  1. NichoPaolucci · · focus · HN ↗
    > OpenAI said it did not believe the industry “has solved alignment and monitoring to a sufficient degree to continue responsibly scaling at maximum speed for much longer.”

    Baffling. To my knowledge, they didn't properly airgap their systems. Keeping the genie in the box seems like 101 to me, and to "miss" that seems awfully fishy. This, among all of the Anthropic news, is an odd convergence.

    Maybe they're being truthful and it really is the end times.

    Maybe they've hit a wall in improvements, but I don't know enough on the topic to speak to that.

    Which is more likely?

    Either way, trying to sift through this can of worms is tiresome. I'm hopeful that this all comes to a head soon, what an exhausting few years it's been...

    1. tim333 · · focus · HN ↗
      My take isn't either. Stopping AI doing bad stuff is a real problem which can likely only be dealt with by trying it out and fixing problems as they arrive. Bit like SpaceX rapid prototyping the rockets - try it out, see what blows up, fix it and try again.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.