Intelligence per Watt: Measuring Intelligence Efficiency of Local AI
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Intelligence per Watt: Measuring Intelligence Efficiency of Local AI
Unofficial Hacker News client; not affiliated with Y Combinator.
polotics · · focus · HN ↗
washadjeffmad · · focus · HN ↗
>Total parameter count governs storage: at FP4, MIXTRAL-8X7B fits within 24 GB GDDR6 (Quadro RTX 6000), and GPT-OSS-120B fits within 128 GB unified memory (Apple M4 Max).