‹ BackHN Continuity

Thread

Qwen3.8 Max now ranked as the best overall model by agentic index

420 points · 270 comments · apitman

  1. syntaxing · · focus · HN ↗
    I am so excited for Qwen 3.8 27B. It’s a shame how slow prefill (~3-400) is on a strix halo but it’s such a good model for agentic tasks.
    1. LoganDark · · focus · HN ↗
      I find that 35B-A3B is much easier to run on my M4 Max (both prefill and generation)
      1. markasoftware · · focus · HN ↗
        It's well known 35b is much faster (on any hardware) and quite a bit dumber
        1. dofm · · focus · HN ↗
          This really very much depends on how you are using it, I think. If you intend to leave it to solve long context problems and write whole prototypes, the 27B is going to be much better.

          But if you are sort of pair-programming with the model, the speed obviously matters and I think then the 35B is acceptably smart, and when it's wrong it'll be wrong much more quickly. It seems very good on SQL and PHP, and I assume on typical JS and Python.

          I would rather work that way, so I hope they do produce a small MoE model.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.