‹ BackHN Continuity

Thread

Alibaba open-sources AI model that can detect cancer and nearly 150 conditions

153 points · 27 comments · yogthos

  1. Zaraif13 · · focus · HN ↗
    I don't agree with the flak Chinese labs get. If it's really that easy to distill and compete with frontier models, why aren't other countries anywhere near this AI race?
    1. felixgallo · · focus · HN ↗
      Because distillation is a friendly term for industrial espionage, and most other countries are not willing to become international pariahs in the eyes of the west.
      1. yogthos · · focus · HN ↗
        meanwhile in the real world <a href="https:&#x2F;&#x2F;www.science.org&#x2F;content&#x2F;article&#x2F;china-tops-world-artificial-intelligence-publications-database-analysis-reveals" rel="nofollow">https:&#x2F;&#x2F;www.science.org&#x2F;content&#x2F;article&#x2F;china-tops-world-art...
        1. felixgallo · · focus · HN ↗
          both things can be true. China&#x27;s gearing up, but they also try to make progress by aggressively distilling Anthropic and OpenAI models, and this is currently where most of their progress comes from.
          1. yogthos · · focus · HN ↗
            People really need to stop parroting this line uncritically. The process takes time because even when you&#x27;re distilling answers, you still need to actually do reinforcement training on the model. And given that Fable and GPT 5.6 just came out there simply hasn&#x27;t been much time to do that. However, models like Kimi also do better than Fable or GPT on a lot of tasks, which means it&#x27;s not just distillation but also difference in architecture. You can watch this talk from Kimi founder to see how Kimi was actually trained and why it performs well. <a href="https:&#x2F;&#x2F;www.youtube.com&#x2F;watch?v=5CkCW1P-g88" rel="nofollow">https:&#x2F;&#x2F;www.youtube.com&#x2F;watch?v=5CkCW1P-g88

            It&#x27;s also absolutely hilarious that people think only Chinese companies use distillation, as if Anthropic or OpenAI are above that or something. Not to mention that they basically ignored copyrights on all the data the siphoned and are now crying that people aren&#x27;t respecting their terms of use.

            Chinese labs have come up with a bunch of genuine innovations: GRPO, auxiliary loss free MoE load balancing, MLA, muon optimizer, and a bunch of other ones. The Deepseek papers are really well written, this isn’t just sneaking a peek at a peer. Anybody who thinks China is simply distilling glorious American models is not engaging with reality.

            1. 4d4m · · focus · HN ↗
              Every model provider distills each other even if for benchmarking. Agree with all your points.
              1. yogthos · · focus · HN ↗
                For sure, everybody distills when they can, it would be stupid not to. I&#x27;m just pointing out that Chinese companies clearly do their own research and innovation just like American companies do. It&#x27;s not that they just wait for American models to drop and then distill them.
                1. 4d4m · · focus · HN ↗
                  Agree completely!
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.