logoalt Hacker News

IanCalyesterday at 9:50 PM1 replyview on HN

How do caches work across models? I would have thought that was very model specific - if not I’ve really misunderstood what’s getting cached.


Replies

armanckeseryesterday at 10:07 PM

I am not sure the author of the comment you are replying to understands that LLM systems have prompt caches