3 ms·
It appears that they do support Prompt Caching: https://inference-docs.cerebras.ai/capabilities/prompt-caching https://inference-docs.cerebras.ai/capabilities/p
by jasongill 23d ago
It appears that they do support Prompt Caching: https://inference-docs.cerebras.ai/capabilities/prompt-caching https://inference-docs.cerebras.ai/capabilities/prompt-cachi...
- abtinf 23d ago> How are cached tokens priced? > There is no additional fee for using prompt caching. Input tokens, whether served from the cache or processed fresh, are billed at the standard input token rate for the respective model. Well, talk about flipping the narrative.
- the_duke 23d agoIt doesn't reduce the price though.
- jasongill 22d agoGood catch, I guess I got lost in the marketing speak of the page!
- deleted 23d ago[deleted]