‹ BackHN Continuity

Thread

MiMo v2.6

1130 points · 483 comments · volf_

  1. rao-v · · focus · HN ↗
    I know we have strong views on what a truly open model is (open weights, open training data, open training code etc.) but I really like how transparent they’ve been about the training of this model.

    The realtime dashboard they shared during training (<a href="https:&#x2F;&#x2F;mimo.xiaomi.com&#x2F;rl&#x2F;" rel="nofollow">https:&#x2F;&#x2F;mimo.xiaomi.com&#x2F;rl&#x2F;) was an incredible learning and teaching tool for me, and they’ve been unusually comprehensive in sharing details about their methodology (check out that tech report - it&#x27;s got lots of clever behind the scene tricks like Google or Deepseek writeups) and benchmark scores (even the stuff they didn’t do well on).

    If you’re releasing an open model going forward, please consider offering the community more of this transparency!

    1. MangoCoffee · · focus · HN ↗
      maybe this is why Dario want to slow down AI development and all the big AI labs in the USA is singing the same song.

      whey they all singing the same tune. it make me question what is their real motives.

      they are afraid of Chinese good enough LLM model killing their margin. we already have story about US companies switch some task to use cheaper Chinese model hosted on Neoclouds.

      1. jwolfe · · focus · HN ↗
        Please explain how putting an upper bound on how good the strongest models can be prevents cheaper less strong models from catching up, rather than enabling it. I do not understand this argument at all.
        1. lelanthran · · focus · HN ↗
          &gt; Please explain how putting an upper bound on how good the strongest models can be prevents cheaper less strong models from catching up, rather than enabling it. I do not understand this argument at all.

          They are not proposing to regulate only the strongest models. They are proposing to regulate all models. If they are already on top, regulation may stop them from proceeding further, but it also stops the cheaper alternatives from catching up.

          If they feel they have reached the asymptote of the curve, then regulation doesn&#x27;t affect them, it affects those who have yet to reach the asymptote.

          1. cogman10 · · focus · HN ↗
            Particularly, the route they seem to want to go is &quot;safety&quot;.

            My guess is that Anthropic and OpenAI will push for &quot;safety&quot; regulations which require byzantine testing that, shocker, Anthropic and OpenAI can pass but the chinese models cannot. The route they&#x27;ll take is import bans and potentially even general bans on products producing or using &quot;unsafe&quot; models.

            They&#x27;ll further likely try and push AI &quot;safety&quot; treaties from the US to other nations to further lock in their lead.

            That&#x27;s why, IMO, we&#x27;ve been seeing so many &quot;OMG, AI will destroy the world and these AI researchers are so scared&quot; articles.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.