‹ BackHN Continuity

Thread

Context Language Models

176 points · 51 comments · emersonmacro

  1. _jayhack_ · · focus · HN ↗
    Letting a model manage its own context is very bitter lesson-pilled

    Biggest challenge is you will get a much lower cache hit rate if you frequently edit the agent's context/prefix, so this can not be implemented efficiently via e.g. the Anthropic API.

    This ^ can be solved in principle but likely requires modifications to the transformer architecture and definitely to serving infrastructure

    See related: &quot;KV Cache Rules Everything Around Me&quot;: <a href="https:&#x2F;&#x2F;www.completeskeptic.com&#x2F;p&#x2F;kv-cache-rules-everything-around" rel="nofollow">https:&#x2F;&#x2F;www.completeskeptic.com&#x2F;p&#x2F;kv-cache-rules-everything-...

    1. Neywiny · · focus · HN ↗
      True facts. Cline does this with auto compact and it&#x27;s so dumb. At the start of the session it reads the plan.md. during compaction, it forgets it. Same with its own output. It&#x27;ll tell me the list of steps to take. By step 4 it needs to compact. So I prompt to do step 4. It spends 15 minutes searching the codebase and comes back that it couldn&#x27;t find such a step. So dumb. So so dumb
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.