‹ BackHN Continuity

Thread

Tokens too cheap to meter

354 points · 227 comments · teoruiz

  1. jetrink · · focus · HN ↗
    > Tokens become cheaper than tool calls

    The author observes that a call to GPT-5.6 Luna is only 4-5 orders of magnitude more expensive than grep, and then predicts that at current rates of progress, calling an LLM will soon be cheaper than a grep. I think this is a good time to invoke Stein's Law: "If something cannot go on forever, it will stop." These efficiency improvements won't continue forever. It's more likely that the per-call cost of high-quality, compiled software like grep will be a lower-bound that LLMs asymptotically approach, rather than a line that they blow past with perpetual exponential progress. (Barring a true breakthrough in something like quantum computing or room-temperature superconductors.)

    1. BearOso · · focus · HN ↗
      4-5 orders of magnitude is huge. Assuming an order of base 10, it's 10000x-100000x. So a call to grep may return in 1s on a typical PC. That means a GPT call takes equivalent energy of 10000-100000 PCs to do the same in 1s. That's a difference that can't be equalized with scaling. It would require a revolutionary breakthrough.

      I also don't understand where the idea that frontier models are getting better efficiency comes from. The results are certainly improving, but that comes from feedback and multiplexing requests, which cost more.

      1. chewbacha · · focus · HN ↗
        Recall the articles claim that efficiency is gaining 2.5 orders of magnitude per year.

        Something seems off about this.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.