logoalt Hacker News

jasongilltoday at 6:46 PM3 repliesview on HN

It appears that they do support Prompt Caching: https://inference-docs.cerebras.ai/capabilities/prompt-cachi...


Replies

the_duketoday at 7:10 PM

It doesn't reduce the price though.

abtinftoday at 7:08 PM

> How are cached tokens priced?

> There is no additional fee for using prompt caching. Input tokens, whether served from the cache or processed fresh, are billed at the standard input token rate for the respective model.

Well, talk about flipping the narrative.

show 1 reply