| ▲ | minimaxir 2 hours ago | |
Per the announcement tweet, BaseTen was the inference provider which has 20% cache cost that is typical: https://www.baseten.co/library/deepseek-v4-flash-0731/ | ||
| ▲ | literallyroy an hour ago | parent [-] | |
Ah thanks. That looks like 10x cost on cache reads vs Deepseek as the provider: https://openrouter.ai/deepseek/deepseek-v4-flash-0731#provid... | ||