‹ BackHN Continuity

Thread

Best LLM for every budget, updated daily

184 points · 113 comments · terryds

  1. Xeoncross · · focus · HN ↗
    If you have a 24-64GB mac, consider running Qwen3.8 27B locally at night. It's a bit slower to run locally, but if you're sleeping it's less of a problem.

    Depending on your memory, you'll need to use the weaker Q4 versions but they still perform well.

    It ranks higher than GPT-5.3 Codex (xhigh) or Claude Opus 4.6 (max) so is great for pairing with <a href="https:&#x2F;&#x2F;github.com&#x2F;kunchenguid&#x2F;gnhf" rel="nofollow">https:&#x2F;&#x2F;github.com&#x2F;kunchenguid&#x2F;gnhf for nightly experimentation, cleanup, or recommendation lists for in the morning.

    1. seanmcdirmid · · focus · HN ↗
      I still haven&#x27;t found a use case for Qwen3.8 27B that Qwen 3.6 35b A3b (MoE) is better at. I can get at most 40 tokens&#x2F;second with 27B, but I get around 90 tokens&#x2F;second with the MoE and it seems to be a more capable model.

      I guess I should still keep experimenting though. Maybe I&#x27;m just not using a dense model correctly.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.