‹ BackHN Continuity

Thread

Qwen3.8 Max now ranked as the best overall model by agentic index

521 points · 327 comments · apitman

  1. onomojo · · focus · HN ↗
    Any benchmark showing Opus 5 as the best just loses credibility for me. Anyone who's actually used Opus 5 daily knows what I'm talking about.
    1. CuriouslyC · · focus · HN ↗
      Ironically, Opus 5 is the most benchmaxxed model I've seen from Anthropic. It is legitimately smart in a lot of ways but it has communication issues, both in terms of how it communicates (all the autism of GPT class models, without the brevity) and how well it catches all the nuance of what you tell it.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.