Remix.run Logo
jasongill 7 hours ago

It appears that they do support Prompt Caching: https://inference-docs.cerebras.ai/capabilities/prompt-cachi...

the_duke 7 hours ago | parent | next [-]

It doesn't reduce the price though.

abtinf 7 hours ago | parent | prev | next [-]

> How are cached tokens priced?

> There is no additional fee for using prompt caching. Input tokens, whether served from the cache or processed fresh, are billed at the standard input token rate for the respective model.

Well, talk about flipping the narrative.

Barbing 7 hours ago | parent | next [-]

heh

Is there a speed increase or is that purely marketing spin on “we might cache on our end but no discount for you”?

lostmsu 6 hours ago | parent [-]

Pure marketing.

qlte an hour ago | parent | prev | next [-]

[dead]

an hour ago | parent | prev [-]
[deleted]
7 hours ago | parent | prev [-]
[deleted]