Remix.run Logo
▲ georgel an hour ago

I'm all in for saving money and _can_ move to using DS directly from them, but maybe I am missing something here:

OpenRouter Pricing:

$0.02/M input tokens $0.60/M output tokens

DeepSeek Pricing (cache miss, off-peak):

$0.15/M Input $0.60/m output

▲girvo an hour ago | parent | next [-]

When 98.5% of my requests are cache hits (according to Pi for the last week), the cache miss price isn’t that important to me, and $0.003-0.006 per 1M input tokens is shockingly cheap.

It’s also the major difference between using DeepSeek directly vs other providers also serving it, though I have not looked lately: it’s possible other providers have matched its cache hit pricing better?

▲georgel an hour ago | parent [-]

Interesting, if the cache hit is that good, I think HN convinced me to toss $20 at DS official, and see how long that lasts.

▲mswphd an hour ago | parent | prev | next [-]

I've heard that certain inference providers may have different quality of caching implementations, so even if the listed numbers are as you say, the practical cache hit % you get might be significantly different/incur significantly different costs.

▲ckdot 17 minutes ago | parent | prev [-]

There’s a big difference in speed & quality between using DeepSeek API directly with DSH vs. DeepSeek in Opencode Go with Opencode CLI. Can’t tell if it’s the provider or the harness - but worth to give it a try.