‹ BackHN Continuity

Thread

Best LLM for every budget, updated daily

184 points · 113 comments · terryds

  1. jrflo · · focus · HN ↗
    Does anyone actually pay API costs out of their own pocket? It's about 10x cheaper to just get a codex or chat gpt subscription, it's so heavily subsidized compared to the API that I'm sure it would be cheaper to use frontier models on a subscription plan rather than paying API prices for deepseek flash.
    1. LeBit · · focus · HN ↗
      I use local LLMs on my Mac Mini.

      Otherwise DeepSeek Flash 4.1 is dirt cheap (other "Flash" models are not that expensive either). I pay (very few dollars) out of my own pocket.

      There are many things where having an API Key is necessary.

      Maybe I’ve missed the boat though: is there now a method to use an api key to access a subscription?

      1. mandeepj · · focus · HN ↗
        > I use local LLMs on my Mac Mini.

        which ones do you use?

        1. LeBit · · focus · HN ↗
          I use the following models:

            - Qwen3.5 9B
            - Qwen3.6 35B A3B
            - Qwen3.8 27B
            - Gemma 4 26B A4B
            - Gemma 31B
            - Muse Glimmer 30B
          
          I have 48G.

          The MacMini is used solely for inference. llama.cpp + llama-swap.

          People are saying local models are crap and serves no purposes.

          I use them to help me spell check, write emails, write text messages, write JIRA tickets, write PR comments, etc.

          I also use them in coding agents to complete different tasks.

          I find them quite useful!

          1. mandeepj · · focus · HN ↗
            Thanks! I have two mac minis: one with 8 gb and the other with 24 gb (left with 17 gb useable ram); I can't run any open model on them. So, ordered a Mac 5 pro with 128 GB. Can't wait for its arrival.
            1. LeBit · · focus · HN ↗
              128G of RAM? I envy you.

              You will be able to run some pretty impressive models locally. Good for you.

              1. mandeepj · · focus · HN ↗
                I'm a small fish.

                Some people here have 640 GB of RAM <a href="https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49621754">https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=49621754

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.