‹ BackHN Continuity

Thread

MiMo v2.6

1130 points · 483 comments · volf_

  1. rao-v · · focus · HN ↗
    I know we have strong views on what a truly open model is (open weights, open training data, open training code etc.) but I really like how transparent they’ve been about the training of this model.

    The realtime dashboard they shared during training (<a href="https:&#x2F;&#x2F;mimo.xiaomi.com&#x2F;rl&#x2F;" rel="nofollow">https:&#x2F;&#x2F;mimo.xiaomi.com&#x2F;rl&#x2F;) was an incredible learning and teaching tool for me, and they’ve been unusually comprehensive in sharing details about their methodology (check out that tech report - it&#x27;s got lots of clever behind the scene tricks like Google or Deepseek writeups) and benchmark scores (even the stuff they didn’t do well on).

    If you’re releasing an open model going forward, please consider offering the community more of this transparency!

    1. figassis · · focus · HN ↗
      Because the world is conditioned to distrust chinese models (pick your reason here), I believe this is critical for them in order to kill any arguments outside the actual merits. They probably spent a lot of time making this call and might pay off on the long run.
      1. embedding-shape · · focus · HN ↗
        Whatever well-founded&#x2F;or not distrust people have in Chinese models, this dashboard proves&#x2F;shows nothing that can make them trust it more or less. It&#x27;s like providing the journalctl logs of your HTTP server on your website and claim this proves NSA isn&#x27;t listening or something.
        1. figassis · · focus · HN ↗
          Theater is often more effective than truth
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.