‹ BackHN Continuity

Thread

How GLM built its own inference infrastructure

411 points · 285 comments · whiteros_e

  1. chung8123 · · focus · HN ↗
    I might be missing something but when I went to their site they are more expensive than Claude. Why would I pick GLM over Claude? Is it they just offer more tokens in their plans?
    1. Bawoosette · · focus · HN ↗
      What are you referring to? Given the audience, my instinct is to assume "plan" refers to the GLM Coding Plans, which are all cheaper than their Anthropic counterparts. As far as I can tell, the API costs are also all cheaper than their roughly equivalently capable Anthropic models.
      1. menaerus · · focus · HN ↗
        Anthropic: 17 USD (pro), 100 USD (max)

        GLM: 80 USD (pro), 168 USD (max) -> with "limited-time event" discount this becomes 56 USD and 117.6 USD

        I also don't understand why are they so much costlier, and I would also like to give it a try.

        1. alexjplant · · focus · HN ↗
          Just use GLM-5.3 Flash via OpenRouter. It's dirt cheap especially relative to how capable it is. While the Z.ai coding plan was a decent deal in the past I always ran into limiting with it and since I use it intermittently for personal projects my usage wasn't always enough to make the math work - the a la carte pricing via OpenRouter makes this a non-issue.

          There's also a new free stealth model available that's more likely than not in the GLM family. This seems to happen every few months for a week or two and represents a good savings opportunity.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.