‹ BackHN Continuity

Thread

MiMo v2.6

1130 points · 483 comments · volf_

  1. XCSme · · focus · HN ↗
    I tried them, but could only test Pro none and Flash none and low, the other ones (medium/high) used way too many tokens and all requests timed out. Not sure if they have a problem with their API, or this model is really token inefficient/basically unusable.
    1. gpugreg · · focus · HN ↗
      The full-response APIs of many providers have been inadequate for a while now because their timeout intervals do not account for lots of thinking. You can use the streaming API to avoid timeouts.
      1. XCSme · · focus · HN ↗
        The timeouts are my self-imposed limits for the test.

        16 minutes per test, which is a lot for simple questions...

        Sometimes they fail because they reason more than their max context window without giving an answer, that's odd too.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.