‹ BackHN Continuity

Thread

Show HN: Run an 80B Qwen in 4.3 GB of RAM on a Mac, and a 35B on an iPhone

310 points · 140 comments · leonickson

  1. myshapeprotocol · · focus · HN ↗
    Achieving high-efficiency model compression to run large models locally on consumer hardware like Macs and iPhones mirrors the architectural goals of decentralized identity.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.