‹ BackHN Continuity

Thread

One month coding with GLM 5.3 Flash

228 points · 180 comments · ThibWeb

  1. monksy · · focus · HN ↗
    GLM5.3-flash has been fantastic for me to make minor fixes in ambigious ways. "Fix x feature, whats going wrong. " It does the job.
    1. girvo · · focus · HN ↗
      Between GLM 5.3 Flash on my legacy Z.ai coding plan, and Qwen 3.8 Flash locally on my DGX Spark-like, I'm barely using my Anthropic/OpenAI subscriptions, likely to cancel them soon.
      1. vardalab · · focus · HN ↗
        It's been slow like molasses on the coding plan. I ended up using it more on fireworks. But DeepSeek is so much cheaper because the cache cost is much better.
        1. girvo · · focus · HN ↗
          I’m kind of lucky that most of the time I don’t have to use it through peak times so it’s not that bad speed wise

          The legacy plan I have is so good as to be basically unlimited usage for my workloads, so I’m kind of stuck with it til they stop renewing it haha

        2. drob518 · · focus · HN ↗
          Cache cost is the primary metric, imo. Most people look at output cost, but you only pay that once. Cache cost you pay every turn and it grows each turn as the context length grows.
        3. pmoriarty · · focus · HN ↗
          The biggest problem with DeepSeek 4.1 Flash is all that it hallucinates a lot.

          [1] - <a href="https:&#x2F;&#x2F;www.youtube.com&#x2F;watch?v=P4dTq4X8bqk" rel="nofollow">https:&#x2F;&#x2F;www.youtube.com&#x2F;watch?v=P4dTq4X8bqk

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.