‹ BackHN Continuity

Thread

The AI Race Just Got Awkward

412 points · 463 comments · allisdust

  1. eggbrain · · focus · HN ↗
    Performance optimizations don't just help the western labs, they also help with running more powerful/useful LLMs locally.

    If local LLMs get "good" enough, people will soon paying for subscriptions to ChatGPT and Claude, which hurts their revenue.

    1. kennywinker · · focus · HN ↗
      The only thing preventing this switch from starting in earnest is the data center buildout monopolizing all current and future GPUs
      1. lumost · · focus · HN ↗
        The margins on NVidia datacenter hardware are ... high. At least one order of magnitude larger than a consumer chip.

        Given the recent deepseekv4.1 advances - how good of a 3B model can we make to run on an iphone natively? is it good enough to match common muse/dot use cases for consumers? the phone is already always on.. no need for a cloud server.

        1. zozbot234 · · focus · HN ↗
          Most phones are not really "always on" in any real sense, a phone on active standby uses very little power and most of it is for its mobile connection. Local AI is best run in a stationary homelab environment, even running it on laptops has its very real problems.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.