‹ BackHN Continuity

Thread

How GLM built its own inference infrastructure

411 points · 285 comments · whiteros_e

  1. embedding-shape · · focus · HN ↗
    I was gonna ask how people found their coding plans, and realized, have they massively ramped up the prices? Seems the middle plan is ~$80/month now, didn't that used to be like $20/month? Cheapest plan is ~$20/month currently.

    They must have hit really hard scaling limits if the prices were hiked so much so quickly.

    1. Daviey · · focus · HN ↗
      I paid $360 annual for Max plan and currently averaging about 1BN tokens a day with their frontier GLM-5.3 model. This was clearly unsustainable for them and they've dropped this package.
      1. world2vec · · focus · HN ↗
        1 billion tokens a day?!! I've done a lot of work these past 2 weeks with GLM-5.3. Like, a lot. And I've just passed 300 million tokens in total.

        Can I ask where are you using all those tokens?

        1. wartywhoa23 · · focus · HN ↗
          Something like this I guess: <a href="https:&#x2F;&#x2F;youtu.be&#x2F;U-Rqv9dOB1U" rel="nofollow">https:&#x2F;&#x2F;youtu.be&#x2F;U-Rqv9dOB1U
          1. absqueued · · focus · HN ↗
            This was such a gem of a video.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.