logoalt Hacker News

VulgarExigencytoday at 12:17 AM1 replyview on HN

The model that is most optimized around token cost is, in fact, Chinese. DeepSeek is astoundingly cheap by default, but if you use it from Reasonix (the harness optimized around its cache), it becomes even cheaper.


Replies

ryeguytoday at 3:57 AM

I keep seeing mention of the cache, what's special about it? All frontier llms have prefix caching, what is special about deepseek's approach?

show 1 reply