‹ BackHN Continuity

Thread

Intelligence per Watt: Measuring Intelligence Efficiency of Local AI

169 points · 65 comments · pythonic_hell

  1. api · · focus · HN ↗
    Unless I misread it, are they saying local GPUs use less energy?

    That’s surprising, almost unbelievable, due to batching. Local is usually not batched.

    1. stymaar · · focus · HN ↗
      Small models are much smaller than frontier models though, which is how they end up consuming less energy despite low batch count. (Though with local models growing strong agentic capabilities, batching becomes a reality with local models as well).
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.