| ▲ | georgel an hour ago | |||||||
I'm all in for saving money and _can_ move to using DS directly from them, but maybe I am missing something here: OpenRouter Pricing: $0.02/M input tokens $0.60/M output tokens DeepSeek Pricing (cache miss, off-peak): $0.15/M Input $0.60/m output | ||||||||
| ▲ | girvo an hour ago | parent | next [-] | |||||||
When 98.5% of my requests are cache hits (according to Pi for the last week), the cache miss price isn’t that important to me, and $0.003-0.006 per 1M input tokens is shockingly cheap. It’s also the major difference between using DeepSeek directly vs other providers also serving it, though I have not looked lately: it’s possible other providers have matched its cache hit pricing better? | ||||||||
| ||||||||
| ▲ | mswphd an hour ago | parent | prev | next [-] | |||||||
I've heard that certain inference providers may have different quality of caching implementations, so even if the listed numbers are as you say, the practical cache hit % you get might be significantly different/incur significantly different costs. | ||||||||
| ▲ | ckdot 17 minutes ago | parent | prev [-] | |||||||
There’s a big difference in speed & quality between using DeepSeek API directly with DSH vs. DeepSeek in Opencode Go with Opencode CLI. Can’t tell if it’s the provider or the harness - but worth to give it a try. | ||||||||