‹ BackHN Continuity

Thread

The AI Race Just Got Awkward

412 points · 463 comments · allisdust

  1. cmiles8 · · focus · HN ↗
    Why is it a problem that the Chinese labs are just distilling down Anthropic’s models? Aren’t Anthropic’s models not just distilling down other people’s work?

    Feels like Anthropic crying do as I say not as I do.

    1. nonethewiser · · focus · HN ↗
      >Aren’t Anthropic’s models not just distilling down other people’s work?

      Can you elaborate on that? I mean my direct answer would be no, of course not. But why do you think frontier models are distilled? I think maybe there is an equivocation over the word “distillation.”

      Frontier labs train on their own pretraining data, human feedback, synthetic data, and research. A distilled model is specifically optimized to reproduce another model's behavior.

      Meanwhile R1-Distill-Qwen-32B was distilled from DeepSeek-R1.

      If you want to say a frontier model is "distilled" from the world's data and R1-Distill-Qwen-32B is distilled from DeepSeek-R1 then you are equivocating two very different things.

      1. nuancebydefault · · focus · HN ↗
        They meant distilling in a more original sense, not per se in the LLM-era meaning of the word sense.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.