‹ BackHN Continuity

Thread

How GLM built its own inference infrastructure

411 points · 285 comments · whiteros_e

  1. embedding-shape · · focus · HN ↗
    I was gonna ask how people found their coding plans, and realized, have they massively ramped up the prices? Seems the middle plan is ~$80/month now, didn't that used to be like $20/month? Cheapest plan is ~$20/month currently.

    They must have hit really hard scaling limits if the prices were hiked so much so quickly.

    1. broodbucket · · focus · HN ↗
      Yeah it went from a great deal to unviable compared to other providers imo. They really need to find a healthy middle ground
      1. lompad · · focus · HN ↗
        It just gives a taste of what we are all going to have to pay soon, once the model providers actually have to make money. And the era of "let's charge a dollar for every 10 dollars running the infra actually costs" is rapidly coming to an end.

        And you can bet GLM is still ridiculously subsidized, just not as ridiculously as Anthropic and OpenAI.

        1. chobbledotcom · · focus · HN ↗
          This isn't true, you can pay for GLM 5.3 from a provider like Neuralwatt or Friendli who have no incentive to subsidize or loss-lead their inference APIs
          1. breakingcups · · focus · HN ↗
            They didn't pay for training
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.