logoalt Hacker News

hn45e7pbijtoday at 1:08 PM1 replyview on HN

Bigger context windows help but they don't remove the need to chunk. Embedding 8K tokens into one vector smears everything, retrieval quality drops even though nothing got truncated.


Replies

btowntoday at 2:43 PM

Are there any good practices on multi-resolution embedding? Like, a strategy where you embed an entire document, and multiple levels of smearing, perhaps going all the way down to 512-token chunks?