Remix.run Logo
SwellJoe 42 minutes ago

DeepSeek is my cheap and cheerful Chinese model of choice for API use. Has been for a while, but now it's Flash instead of Pro. Even cheaper, and now better then Pro. I feel like most of the major Chinese models are benchmaxxed, they have weird quirks every time I use them (Qwen 3.8 Max doesn't check its work and leaves stuff broken, doesn't write tests unless prompted, etc., Kimi ends up being quite expensive and rarely better than GPT Sol or Opus 5), while DeepSeek models seem to be generally as good as the benchmarks indicate: Not the best, but stronger across the board than any model within an order of magnitude of its price.

eli 34 minutes ago | parent [-]

Qwen 3.8 Max is very strong at troubleshooting and code review.

SwellJoe 6 minutes ago | parent [-]

I'll grant it's very thorough when assigned a troubleshooting task. I'm not as confident of it's code review though it is very good at security vulnerability auditing, and isn't hobbled for that work like Fable, and even Opus refuses some work in that area now.