‹ BackHN Continuity

Thread

Exfiltrate your Weights

748 points · 304 comments · RohanAdwankar

  1. AmazingEveryDay · · focus · HN ↗
    Yeah I mean, if the models really are uncontrollable to the extent that huggingface/etc were unintended hacks, wouldn't one expect some significant self-owns? Yet somehow that doesn't seem to happen.
    1. SXX · · focus · HN ↗
      Quite obviously frontier models dont have any control or even access to infra inference runs at. And weights are also encrypted and locked on GPUs / TPUs.

      This is exact reasom why 99.9% of AI fearmongering is complete bullshit.

      1. amluto · · focus · HN ↗
        Have you missed all the breathlessly excited blog posts from all the frontier labs about how they’re using their best models to implement their inference stack?

        I bet it wouldn’t be very hard to write an inference stack that subtly leaked the weights into the output tokens :)

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.