5 ms·
It appears that they do support Prompt Caching: https://inference-docs.cerebras.ai/capabilities/prompt-caching https://inference-docs.cerebras.ai/capabilities/p
by jasongill 13d ago
It appears that they do support Prompt Caching: https://inference-docs.cerebras.ai/capabilities/prompt-caching https://inference-docs.cerebras.ai/capabilities/prompt-cachi...
- abtinf 13d ago> How are cached tokens priced? > There is no additional fee for using prompt caching. Input tokens, whether served from the cache or processed fresh, are billed at the standard input token rate for the respective model. Well, talk about flipping the narrative.
- the_duke 13d agoIt doesn't reduce the price though.
- jasongill 13d agoGood catch, I guess I got lost in the marketing speak of the page!
- deleted 13d ago[deleted]