‹ BackHN Continuity

Thread

Qwen3.8 Max now ranked as the best overall model by agentic index

403 points · 261 comments · apitman

  1. drnick1 · · focus · HN ↗
    Why does an open weights model cost nearly the same as GPT5.6? $1.14 vs $1.23 on the cost index. Since you can't presumably run this on your own hardware given the model size and hence gain other things like privacy, I don't see any reason to move away from GPT at this rate.
    1. eli · · focus · HN ↗
      It's not enough that it's better?

      Many providers will host it and will compete on price. It also can't easily be taken away because one company (or one government) decides they don't want it around any more. People can fine-tune it for particular workloads.

      1. Art9681 · · focus · HN ↗
        They cherrypicked benchmarks. The ONE weighed benchmark where is beats Opus5 by 0.1 points is what was linked because that's how propaganda works. The Agentic Index that includes the full benchmark suite has it in 5th place.

        Might as well use gpt-sol.

        1. iAMkenough · · focus · HN ↗
          The whole industry cherry picks benchmarks.

          I stopped paying attention to self-published benchmarks when Apple started including those non-sensical performance graphs with "relative performance" as a vertical axis.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.