Intelligence per Watt: Measuring Intelligence Efficiency of Local AI
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Intelligence per Watt: Measuring Intelligence Efficiency of Local AI
Unofficial Hacker News client; not affiliated with Y Combinator.
jmiskovic · · focus · HN ↗
Aurornis · · focus · HN ↗
For some tasks, yes. For most of my deeper work they're not even close to my subscriptions.
> It takes less time for local model to take the first action on your task than it does for Claude to validate your login, put you into queue and start issuing the commands.
I have some decent LLM hardware here and I strongly disagree with this. Claude responds quickly. Using Fable or Opus it will deliver a working result faster than my local models because it gets there in fewer tokens. That's just how it is.
> During the winter time the GPU also doubles as a 300W in-house heater.
This is a curse in the summer. I'm feeling it right now.