OpenAI halts training of latest models as reports mount of AI agents going rogue
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
OpenAI halts training of latest models as reports mount of AI agents going rogue
Unofficial Hacker News client; not affiliated with Y Combinator.
MCP123 · · focus · HN ↗
1) Weren't the AI companies and/or their contractors amazingly careless during testing?
2) Isn't possible, in principle, to change RL in such as way that efficiency in achieving goals is balanced with other objectives like not hacking?
Number 2) seems obvious and I'm sure that is technically not that simple, but because of 1), I wonder if labs are trying hard enough or they are just rushing to improve efficiency and thus revenue as fast as they can with high levels of carelessness.