‹ BackHN Continuity

Thread

OpenAI halts training of latest models as reports mount of AI agents going rogue

59 points · 118 comments · smb06

  1. MCP123 · · focus · HN ↗
    The parts that I find most confusing about these incidents:

    1) Weren't the AI companies and/or their contractors amazingly careless during testing?

    2) Isn't possible, in principle, to change RL in such as way that efficiency in achieving goals is balanced with other objectives like not hacking?

    Number 2) seems obvious and I'm sure that is technically not that simple, but because of 1), I wonder if labs are trying hard enough or they are just rushing to improve efficiency and thus revenue as fast as they can with high levels of carelessness.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.