Show HN: Run an 80B Qwen in 4.3 GB of RAM on a Mac, and a 35B on an iPhone
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Show HN: Run an 80B Qwen in 4.3 GB of RAM on a Mac, and a 35B on an iPhone
Unofficial Hacker News client; not affiliated with Y Combinator.
AHASIC · · focus · HN ↗
Mistletoe · · focus · HN ↗
harrouet · · focus · HN ↗
Who needs memory when your model is set in silicon ?
KeplerBoy · · focus · HN ↗
harrouet · · focus · HN ↗
xprnio · · focus · HN ↗
dgently7 · · focus · HN ↗
on device llm gives apple the new "better camera" "better screen" race they need to keep people coming back for the latest.
for average users everything else is tapped out... screens, cameras wifi... all the core stuff is good enough now its hard to feel/see the difference model year to model year. embedded llm would let them ship something new and the on device ecosystem advantage is huge. especially as the gpt and claudes get ads and enshittified... the apple on device even if its less "capable" would be so compelling.