Remix.run Logo
minimaxir 2 hours ago

Per the announcement tweet, BaseTen was the inference provider which has 20% cache cost that is typical: https://www.baseten.co/library/deepseek-v4-flash-0731/

literallyroy an hour ago | parent [-]

Ah thanks. That looks like 10x cost on cache reads vs Deepseek as the provider: https://openrouter.ai/deepseek/deepseek-v4-flash-0731#provid...