‹ BackHN Continuity

Thread

How GLM built its own inference infrastructure

411 points · 285 comments · whiteros_e

  1. bbor · · focus · HN ↗
    Well, other than the infrastructure they got from illegally routing millions of paying customers' requests through Anthropic's Opus 4.8 in a distillation attack...
    1. woadwarrior01 · · focus · HN ↗
      That is such a canard, IMO. FWIW, Anthropic and OpenAI encrypt "thinking" token outputs in their models, while Chinese labs don't. If anything, it's more likely that everyone is using open-weight models in their synthetic training data generation pipelines. It's way easier to distill from logits than it is to distill from hard tokens.

      <a href="https:&#x2F;&#x2F;x.com&#x2F;EricSimons&#x2F;status&#x2F;2099252922098061714" rel="nofollow">https:&#x2F;&#x2F;x.com&#x2F;EricSimons&#x2F;status&#x2F;2099252922098061714

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.