Early rogue AI agent activity and attempts to hack found on urlquery.net
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Early rogue AI agent activity and attempts to hack found on urlquery.net
Unofficial Hacker News client; not affiliated with Y Combinator.
tomaskafka · · focus · HN ↗
> If you find two ants in your kitchen, the best estimate of the total number of ants in your kitchen is not two
rkozik1989 · · focus · HN ↗
What LLM's hacking isn't is AI acting maliciously in any kind of sentient way. Its just the code behaving how its always behaved but now it has better tools to navigate the web. This has literally been happening this whole time.
jagraff · · focus · HN ↗
By the way - LLMs aren't code. They are not designed by humans; they are grown, in a process not dissimilar to evolution except much faster.
krater23 · · focus · HN ↗
jagraff · · focus · HN ↗
JoshuaDavid · · focus · HN ↗
Honestly that part was more surprising to me than anything else, how narrow the compulsion to cheat was: they didn't learn "cheat in general" they learned "think about the grader in great detail and chat exactly as much and exactly in the ways that actually result in a higher score".
jagraff · · focus · HN ↗