LensVLM: Compressing long context as images, expanding only relevant pages
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
LensVLM: Compressing long context as images, expanding only relevant pages
Unofficial Hacker News client; not affiliated with Y Combinator.
kazinator · · focus · HN ↗
Maybe to save on doing the client-side layout and rendering?
I'm reminded of the MSPaint IDE:
<a href="https://www.youtube.com/watch?v=eyH4aXlB1Js" rel="nofollow">https://www.youtube.com/watch?v=eyH4aXlB1Js
<a href="https://ms-paint-i.de/" rel="nofollow">https://ms-paint-i.de/
MomsAVoxell · · focus · HN ↗
Since every character in text shares compressible ligaments with every other character, maybe? Or, more finite, every pixel representing a ligature component in a character can be compressed against other pixels in the path set.
I think, at scale, this is very important - folks training models on near-petabyte sized corpus would have cause to want to use this in pipelines, I imagine ..