AI companies in race to demonstrate their model most threatening to humanity
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
AI companies in race to demonstrate their model most threatening to humanity
Unofficial Hacker News client; not affiliated with Y Combinator.
ACCount39 · · focus · HN ↗
OpenAI, for example, thought their sandboxes were good enough. As their AIs got more and more advanced, they kept proving them wrong - sandbox after sandbox.
And that's today's AI problems. AI capabilities are still improving - if there's a limit to that, we are yet to find it. Coupled with how willing today's AIs are to break the rules and resort to "hack the world" in their problem solving? Very concerning.
nicce · · focus · HN ↗
What I have been reading, was that their sandboxes were so poor that it was pure negligence. I am still waiting to see if some external and neutral cybersecurity company with high reputation would audit their sandboxes and how they are being used.
simianwords · · focus · HN ↗
How were the sandboxes poor?
dns_snek · · focus · HN ↗
simianwords · · focus · HN ↗
dns_snek · · focus · HN ↗
simianwords · · focus · HN ↗
dns_snek · · focus · HN ↗
There was no real isolation because a part of the system that doesn't provide any isolation guarantees was bridged to the internet. The next version of Artifactory, which you definitely wouldn't audit before you rolled it out, could simply add a public API that sends requests out to the internet.
Such an innocent upstream change would be equally catastrophic for your security model, which should demonstrate why it's negligent to rely on undefined behavior to enforce your security policies.
This should frankly be obvious any operator entrusted with running dangerous and possibly malicious code. Even if you don't know what you're doing, any LLM would tell you that this is a really bad idea if you simply asked. Don't rely on spacebar heating [1] to keep humanity alive.
[1] <a href="https://xkcd.com/1172/" rel="nofollow">https://xkcd.com/1172/