‹ BackHN Continuity

Thread

Claude Opus 5.5 Intelligence, Performance and Price Analysis (Max)

333 points · 106 comments · theanonymousone

  1. hglaser · · focus · HN ↗
    Half the cost per task compared to Opus 5, comparing high effort to high effort. That's just really nice.

    Edit: <a href="https:&#x2F;&#x2F;artificialanalysis.ai&#x2F;models&#x2F;claude-opus-5-5?models=claude-opus-5-5-high%2Cclaude-opus-5-high#price-cost" rel="nofollow">https:&#x2F;&#x2F;artificialanalysis.ai&#x2F;models&#x2F;claude-opus-5-5?models=...

    1. user43928 · · focus · HN ↗
      Astra High is slightly cheaper at $1.73 vs $1.82 for Opus 5.5
      1. onlyrealcuzzo · · focus · HN ↗
        The UI&#x2F;UX seems impressively bad. DeepSWE&#x27;s cost curve has a better, more obvious way to sort by only the top level of reasoning to avoid 80% of the graph just being the same 3-5 models at their 8 different reasoning levels...

        It&#x27;s also less clear what a lot of their metrics mean. Does Cost per Task include only things that can be verified to work and passed? As best I can tell, it does not.

        I&#x27;m less concerned if one model&#x27;s cost per task is $0.10 and another model&#x27;s cost is $1.50 if the $0.10 task got it right 1% of the time and the $1.50 model got it right 66% of the time.

        An equalized &#x2F; weighted cost&#x2F;time per task is much more valuable - being massively penalized for taking a lot of time and ultimately not passing when OTHER models did pass.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.