logoalt Hacker News

Invictus0 • today at 1:43 PM • 6 replies • view on HN

Can someone please explain to me how the recent llm cache breakthrough doesn't alleviate this memory shortage issue? https://intl.cloud.baidu.com/en/article/8937874


Replies

selbyk • today at 1:55 PM

Why use less memory when you can just have more LLM?

➕ show 1 reply
Grimeton • today at 1:54 PM

I have this wild theory that this has not just todo with AI but with weapons production overall.

➕ show 1 reply
gortok • today at 1:45 PM

Not everyone is using baidu?

linuxftw • today at 2:00 PM

If training and inference is hardware constrained, and you can train and server better and bigger models on the same hardware with memory optimizations, that's exactly what I would expect companies to do.

s0ss • today at 2:01 PM

Demand outstrips supply. Efficient software is great, but it doesn’t directly resolve insufficient supply of hardware.