logoalt Hacker News

seaurchinzeetoday at 6:17 PM1 replyview on HN

"Cache reads now cost 75% less, or $0.25 per million tokens." For me, at a typical 95% cache hit rate, I think my optimal context window size before autocompaction goes from ~200K to ~400K tokens. Great for longer horizon tasks.


Replies

cute_boitoday at 6:22 PM

looks like it is only for api.....

show 3 replies