‹ BackHN Continuity

Thread

Better Vector Search for Long Documents: Chunking Inside Manticore Search

80 points · 14 comments · GloriaVinogrado

  1. hn45e7pbij · · focus · HN ↗
    Bigger context windows help but they don't remove the need to chunk. Embedding 8K tokens into one vector smears everything, retrieval quality drops even though nothing got truncated.
    1. btown · · focus · HN ↗
      Are there any good practices on multi-resolution embedding? Like, a strategy where you embed an entire document, and multiple levels of smearing, perhaps going all the way down to 512-token chunks?
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.