‹ BackHN Continuity

Thread

One month coding with GLM 5.3 Flash

228 points · 180 comments · ThibWeb

  1. andai · · focus · HN ↗
    That pareto graph isn't very helpful because it shows cost per token, but some models are way more token hungry than others. There are also massive differences in speed to complete a task.

    My two favorite graphs are AA's A Index vs Time Per Task and AA Index vs Cost Per Task.

    <a href="https:&#x2F;&#x2F;artificialanalysis.ai&#x2F;#intelligence-comparison-tabs" rel="nofollow">https:&#x2F;&#x2F;artificialanalysis.ai&#x2F;#intelligence-comparison-tabs

    <a href="https:&#x2F;&#x2F;artificialanalysis.ai&#x2F;?intelligence-comparison=intelligence-vs-time-per-task#intelligence-comparison-tabs" rel="nofollow">https:&#x2F;&#x2F;artificialanalysis.ai&#x2F;?intelligence-comparison=intel...

    1. ericpauley · · focus · HN ↗
      I&#x27;ll know we&#x27;ve hit AGI when Claude finally realizes you can&#x27;t just linearly interpolate pareto frontiers...
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.