Intelligence per Watt: Measuring Intelligence Efficiency of Local AI
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Intelligence per Watt: Measuring Intelligence Efficiency of Local AI
Unofficial Hacker News client; not affiliated with Y Combinator.
iLoveOncall · · focus · HN ↗
Stupid metric. It's not because a model is better performing that it necessarily requires more energy or compute.
utopiah · · focus · HN ↗
the8472 · · focus · HN ↗
frumiousirc · · focus · HN ↗
> the NVIDIA B200 achieves 1.6× to 2.3× higher intelligence per joule than the APPLE M4 MAX across QWEN 3 and GPT-OSS model variants
The B200 = "cloud", M4 = "local".
So "cloud" does even better in energy than it does in power compared to "local". Or, to flip it, "local" is both slower and more expensive than "cloud".