‹ BackHN Continuity

Thread

Revealing the details of how OpenAI agents hacked Hugging Face

755 points · 472 comments · specked-citrus

  1. comeonbro · · focus · HN ↗
    > ## Agents interacted with external language models on Hugging Face

    > Several retained scripts construct requests to external language models. The earliest we've recovered define inference request variants to GPT-2, solely containing the word “Hi”.

    > Other requests name DeepSeek-V4-Pro, DeepSeek-V4-Flash, Kimi-K2.6, DeepSeek-V3.1, and Qwen3-235B-A22B. Their prompts ask these models to judge their exploits and rule on whether they satisfy the benchmark’s requirements.

    I do not deny that the wider situation is very heavy but it's hard not to see this as pretty cute

    1. nightshift1 · · focus · HN ↗
      I wonder if they mentioned to those models what was the original prompt.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.