‹ BackHN Continuity

Thread

Best LLM for every budget, updated daily

184 points · 113 comments · terryds

  1. Xeoncross · · focus · HN ↗
    If you have a 24-64GB mac, consider running Qwen3.8 27B locally at night. It's a bit slower to run locally, but if you're sleeping it's less of a problem.

    Depending on your memory, you'll need to use the weaker Q4 versions but they still perform well.

    It ranks higher than GPT-5.3 Codex (xhigh) or Claude Opus 4.6 (max) so is great for pairing with <a href="https:&#x2F;&#x2F;github.com&#x2F;kunchenguid&#x2F;gnhf" rel="nofollow">https:&#x2F;&#x2F;github.com&#x2F;kunchenguid&#x2F;gnhf for nightly experimentation, cleanup, or recommendation lists for in the morning.

    1. jszymborski · · focus · HN ↗
      It&#x27;s _so_ good, I no longer bother with Sonnet and use it locally for everything.

      Consider bumping reasoning down to Medium as a default though, I agree with simonw it over thinks <a href="https:&#x2F;&#x2F;simonwillison.net&#x2F;2026&#x2F;Aug&#x2F;16&#x2F;qwen-38-27b&#x2F;" rel="nofollow">https:&#x2F;&#x2F;simonwillison.net&#x2F;2026&#x2F;Aug&#x2F;16&#x2F;qwen-38-27b&#x2F;

      1. felineflock · · focus · HN ↗
        Isn&#x27;t there a way to set up a thinking budget so it automatically tells the model &quot;Conclude your reasoning now and provide the final answer.&quot; ?
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.