Remix.run Logo
gymbeaux 3 hours ago

Kimi and DeepSeek are impressive for their size but still run slow on any hardware you or I would have. If you’re proposing we run those via a cloud service, I don’t see a reason to do that when it’s an inferior model and I still have to pay per token for it.

alexjplant 3 hours ago | parent | next [-]

OpenCode Go provides a lot of usage for these models for the paltry sum of $10/month. Z.ai's coding plan provides a single-digit multiple of Claude Code's usage for a similar price and performance level. Kimi and DeepSeek models are hundreds of billions of parameters (or, in K3's case, >1T). Many of these models have Opus-level benchmarks and, as I pointed out previously, often practically outperform Anthropic models because they're more consistent.

throw1234567891 2 hours ago | parent | prev [-]

You should run the maths once. Those tokens cost you much more than the hardware would. But yeah, CAPEX vs OPEX something something.