‹ BackHN Continuity

Thread

Step 5 Preview: Advancing the Pareto Frontier

141 points · 33 comments · nateb2022

  1. Jacques2Marais · · focus · HN ↗
    For anyone else looking for the pricing: <a href="https:&#x2F;&#x2F;platform.stepfun.ai&#x2F;docs&#x2F;en&#x2F;guides&#x2F;pricing&#x2F;details#pricing-for-multimodal-reasoning-models" rel="nofollow">https:&#x2F;&#x2F;platform.stepfun.ai&#x2F;docs&#x2F;en&#x2F;guides&#x2F;pricing&#x2F;details#p...
    1. sieve · · focus · HN ↗
      I regularly hit 200-300M cached reads every day on some of the models I use. It has exceeded 7-800M on a couple of occasions. At $0.04&#x2F;M, that is $8-12 per day only for cached reads.
      1. ignoramous · · focus · HN ↗
        &gt; At $0.04&#x2F;M

        Unless you meant step-3.7-flash, the input cache hits are $0.05 per mil for step-5-preview.

        &gt; $8-12 per day only for cached reads

        Pretty decent &quot;API&quot; rates for ~500M+ tokens on Step Fun 5, a Kimi K3 &#x2F; GLM 5.3 level model?

        Their &quot;Step Plan&quot; is ridiculous, by comparison: ~$60 usage on $6.99&#x2F;mo; ~$220 on $9.99&#x2F;mo. <a href="https:&#x2F;&#x2F;platform.stepfun.ai&#x2F;docs&#x2F;en&#x2F;step-plan&#x2F;overview" rel="nofollow">https:&#x2F;&#x2F;platform.stepfun.ai&#x2F;docs&#x2F;en&#x2F;step-plan&#x2F;overview

        1. sieve · · focus · HN ↗
          Yes, I meant the Flash version.

          I have used Kimi 2.5 and GLM 5.3 (&amp; 5.3 Flash). Do not need them for what I do outside of spec hardening (basically, a lot of chatting).

          I tend to know exactly what I want and most of the weaker models are enough to get me there. I have mainly been using MiMo, DeepSeek V4 Flash and MuseSpark Contributor over the last month or so.

Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.