‹ BackHN Continuity

Thread

Mercury 2.5 LLM hits 770 tokens per second

151 points · 92 comments · Retro_Dev

  1. freakynit · · focus · HN ↗
    I have tried using Mercury 2.5 for a lot of my tasks.. but this model just isn't there. It seems to be on par with any 14B model at max. Even GPT-OSS-20B performs way better than this in my own attempts to use it.

    I really really wanted to use this because it offers incredible speeds and pricing combinations. But nop.. I still am not using it.. not even for basic tasks.

    1. nostrebored · · focus · HN ↗
      Try Celeris-magnus-1. We get similar speeds and it’s much closer to qwen 27B dense models.
      1. freakynit · · focus · HN ↗
        Didn't knew about this. Thanks.. but, according artificialanalysis.ai, it's intelligence is just about like a 14B model (mistral 3 14b)
        1. nostrebored · · focus · HN ↗
          depends on your use case I think. for small, transactional tasks it is quite strong and fast.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.