‹ BackHN Continuity

Thread

Cache-to-Cache: Direct Semantic Communication Between LLMs (2025)

109 points · 22 comments · rochansinha

  1. gavinray · · focus · HN ↗
    A few months ago I asked why semantic representation rather than text wasn't used, since natural language seems quite a lossy representation for semantic concepts:

    <a href="https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=47195212">https:&#x2F;&#x2F;news.ycombinator.com&#x2F;item?id=47195212

    I wouldn&#x27;t have thought to use it for LLM-to-LLM communication, though

    1. msdz · · focus · HN ↗
      My guess is that if you go with something other than readable text as the “thought layer”, observability becomes impossible.

      Which is not necessarily something the humans training a model would want re&#x2F; alignment.

      1. TeMPOraL · · focus · HN ↗
        Indeed this is one of the key problems here. If models start doing CoT in &quot;neuralese&quot;, or use it when talking with each other, we lose what little observability we have.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.