| ▲ | gymbeaux 3 hours ago | |
Kimi and DeepSeek are impressive for their size but still run slow on any hardware you or I would have. If you’re proposing we run those via a cloud service, I don’t see a reason to do that when it’s an inferior model and I still have to pay per token for it. | ||
| ▲ | alexjplant 3 hours ago | parent | next [-] | |
OpenCode Go provides a lot of usage for these models for the paltry sum of $10/month. Z.ai's coding plan provides a single-digit multiple of Claude Code's usage for a similar price and performance level. Kimi and DeepSeek models are hundreds of billions of parameters (or, in K3's case, >1T). Many of these models have Opus-level benchmarks and, as I pointed out previously, often practically outperform Anthropic models because they're more consistent. | ||
| ▲ | throw1234567891 2 hours ago | parent | prev [-] | |
You should run the maths once. Those tokens cost you much more than the hardware would. But yeah, CAPEX vs OPEX something something. | ||