The margins on NVidia datacenter hardware are ... high. At least one order of magnitude larger than a consumer chip.
Given the recent deepseekv4.1 advances - how good of a 3B model can we make to run on an iphone natively? is it good enough to match common muse/dot use cases for consumers? the phone is already always on.. no need for a cloud server.
eggbrain · · focus · HN ↗
If local LLMs get "good" enough, people will soon paying for subscriptions to ChatGPT and Claude, which hurts their revenue.
kennywinker · · focus · HN ↗
lumost · · focus · HN ↗
Given the recent deepseekv4.1 advances - how good of a 3B model can we make to run on an iphone natively? is it good enough to match common muse/dot use cases for consumers? the phone is already always on.. no need for a cloud server.
christkv · · focus · HN ↗