‹ BackHN Continuity

Thread

Qwen3.8 Max now ranked as the best overall model by agentic index

295 points · 178 comments · apitman

  1. embedding-shape · · focus · HN ↗
    Strange that the page <a href="https:&#x2F;&#x2F;artificialanalysis.ai&#x2F;agents&#x2F;coding-agents" rel="nofollow">https:&#x2F;&#x2F;artificialanalysis.ai&#x2F;agents&#x2F;coding-agents doesn&#x27;t even mention &quot;Qwen&quot; once if it&#x27;s now the &quot;best&quot; according to one of their one index?
    1. artemisart · · focus · HN ↗
      They didn&#x27;t run all benchmarks. It&#x27;s the best in AA agentic index (GDPval-AA v2, ³-Banking) but not coding index (DeepSWE which is missing, Terminal-Bench v2.1 they have 81% vs 90% for Sol, SWE-Atlas-QnA missing).
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.