‹ BackHN Continuity

Thread

Best LLM for every budget, updated daily

184 points · 113 comments · terryds

  1. Xeoncross · · focus · HN ↗
    If you have a 24-64GB mac, consider running Qwen3.8 27B locally at night. It's a bit slower to run locally, but if you're sleeping it's less of a problem.

    Depending on your memory, you'll need to use the weaker Q4 versions but they still perform well.

    It ranks higher than GPT-5.3 Codex (xhigh) or Claude Opus 4.6 (max) so is great for pairing with <a href="https:&#x2F;&#x2F;github.com&#x2F;kunchenguid&#x2F;gnhf" rel="nofollow">https:&#x2F;&#x2F;github.com&#x2F;kunchenguid&#x2F;gnhf for nightly experimentation, cleanup, or recommendation lists for in the morning.

    1. ghilston · · focus · HN ↗
      What would you personally recommend for those that have 128 GB?
      1. seanmcdirmid · · focus · HN ↗
        if you have 128 GB, you could use Qwen Flash Next at some reasonable quant, with the new SSD hack for only keeping some of the model resident in memory.
      2. nolok · · focus · HN ↗
        I have a Ryzen AI Max+ 395 with 128 GB for running those sort of &quot;background task&quot;, and my sweet spot is currently Qwen3.8 Flash-Next IQ4 at 96 GB.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.