‹ BackHN Continuity

Thread

One month coding with GLM 5.3 Flash

228 points · 180 comments · ThibWeb

  1. sheepscreek · · focus · HN ↗
    > 450M tokens / $150 / 5kWh

    Makes me appreciate my ChatGPT subscription. I’ve had multiple days between 1B-2B tokens (now less so, models have indeed become token efficient) and regularly in the > 100M range. Even then, $150 sounds excessive. I wonder if their cache is getting nuked for some reason, or maybe they decide to use Cerebras that doesn’t subsidize cached tokens.

    1. sheepscreek · · focus · HN ↗
      EDIT: Proof (can't edit the original comment)

      <a href="https:&#x2F;&#x2F;imgur.com&#x2F;a&#x2F;mSXGzzJ" rel="nofollow">https:&#x2F;&#x2F;imgur.com&#x2F;a&#x2F;mSXGzzJ

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.