‹ BackHN Continuity

Thread

MiMo v2.6

1130 points · 483 comments · volf_

  1. rao-v · · focus · HN ↗
    I know we have strong views on what a truly open model is (open weights, open training data, open training code etc.) but I really like how transparent they’ve been about the training of this model.

    The realtime dashboard they shared during training (<a href="https:&#x2F;&#x2F;mimo.xiaomi.com&#x2F;rl&#x2F;" rel="nofollow">https:&#x2F;&#x2F;mimo.xiaomi.com&#x2F;rl&#x2F;) was an incredible learning and teaching tool for me, and they’ve been unusually comprehensive in sharing details about their methodology (check out that tech report - it&#x27;s got lots of clever behind the scene tricks like Google or Deepseek writeups) and benchmark scores (even the stuff they didn’t do well on).

    If you’re releasing an open model going forward, please consider offering the community more of this transparency!

    1. bicepjai · · focus · HN ↗
      I was absolutely mind blown when I saw how they were publishing that training dashboard while US models publish 100s of pages of reports (just provide a &quot;copy as MD&quot; button, folks, in the future). I was thinking about doing something similar but did not know how to show it, and this is a perfect example for someone who wants to show whatever they are training, for me it was local training on a consumer GPU.

      My dream is to see this like a dashboard for a model trained across distributed machines, like Bitcoin mining, where minted coins are given to people whose machines were used for training. I don&#x27;t know if they are worth it, but bragging rights alone, like a tag they can put on a website or social media, will be good enough for me.

      1. sally_glance · · focus · HN ↗
        Not an expert on this but I think the RL runs need to work sequentially? I wonder what the opportunities for distributed execution would be... Maybe parallelizing the benchmark task or inference
        1. whimsicalism · · focus · HN ↗
          i think you’re wrong - async rollouts very much standard and i would be shocked if benchmark evals were not done in parallel
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.