‹ BackHN Continuity

Thread

DeepSeek-v4.1 Flash: Pushing the Limits of KV Cache Compression

131 points · 10 comments · mfiguiere

  1. mmastrac · · focus · HN ↗
    I've been working with an automatic incremental context compactor enabled and it's been surprisingly helpful. It was particularly effective with DS41f - I think I was running at an effective session length of 5M, with the model running around 300k-400k and it was holding on both speed and intelligence.

    TBH I also ran the 400tok/s preview and that was just nuts. I just let the thing compact over and over over the course of a day attacking a couple of tough problems

    1. nchmy · · focus · HN ↗
      Can you share a link to thr automatic compactor? I've been noticing that when I get to around 60% context window, the cache will simply break and suddenly I've paid 50x more than expected. The only solution seems to be to compact or start a new session.
      1. mmastrac · · focus · HN ↗
        Send me an email- it's not public just yet
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.