‹ BackHN Continuity

Thread

Qwen3.8 Max now ranked as the best overall model by agentic index

420 points · 270 comments · apitman

  1. drnick1 · · focus · HN ↗
    Why does an open weights model cost nearly the same as GPT5.6? $1.14 vs $1.23 on the cost index. Since you can't presumably run this on your own hardware given the model size and hence gain other things like privacy, I don't see any reason to move away from GPT at this rate.
    1. eli · · focus · HN ↗
      It's not enough that it's better?

      Many providers will host it and will compete on price. It also can't easily be taken away because one company (or one government) decides they don't want it around any more. People can fine-tune it for particular workloads.

      1. drnick1 · · focus · HN ↗
        > It's not enough that it's better?

        It's barely better, and barely cheaper, not really enough to challenge the status quo IMO. Half the price for basically the same performance would be a much stronger value proposition.

        1. ux266478 · · focus · HN ↗
          What status quo? Just look at Openrouter&#x27;s rankings: <a href="https:&#x2F;&#x2F;openrouter.ai&#x2F;rankings" rel="nofollow">https:&#x2F;&#x2F;openrouter.ai&#x2F;rankings

          Things change radically month to month. Nobody is remotely close to capturing the market or having any kind of stability over time. People move around quite a lot, often to sidegrade within a generation. Just playing fly on the wall with discourse would be enough to tell you all of this, even without the data to back it up.

          1. eli · · focus · HN ↗
            That&#x27;s got a significant selection bias. Claude and ChatGPT and Gemini and other subs do not go through openrouter.
            1. ux266478 · · focus · HN ↗
              Not really, because that&#x27;s not a unique aspect of any of those. It&#x27;s true of all subscription services (that I&#x27;m aware of), as well as all of the free models. The selection bias primarily will be against models which be an outlier in the difference between openrouter users and total users, which is a much harder position to argue for any given company except for maybe Twitter.

              You can argue there&#x27;s a selection bias that openrouter users are less likely to display model loyalty, but it would still be a visible confounding factor if it was a statistically significant behavior. And it&#x27;s not. Nor is there a visibly meaningful indication that people don&#x27;t sidegrade between models. With every single data set, you&#x27;re going to see that. You&#x27;re also going to see it reflected in discourse, as I mentioned. Fact of the matter is there isn&#x27;t a status quo in AI any more than there&#x27;s a status quo in cars.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.