Remix.run Logo
▲ jacquesm 3 hours ago

If DS4.1 impresses you I would be really interested to see your comparison to GLM 5.3. I switched from the one to the other and even if GLM 5.3 is a bit slower I don't think I'll be going back.

▲badatnames 2 hours ago | parent | next [-]

DS4 (not 4.1) crossed my dont-care threshold and I genuinely stopped paying attention to new models. I'd love to try GLM 5.3 but I just don't see any point in spending the effort any more. I can get passable intelligence for a bargain price either direct from China or from a ZDR EU provider for a small markup. Paying 10x more will not make me 10x happier, it's unlikely to make me even 1.1x happier now I've got some intuition for the natural limits of these models.

I don't even bother checking how much I spent on API any more, its well under $30 over the past 2 months despite daily constant use. Who even needs a subscription at these numbers?

▲ctolsen 3 hours ago | parent | prev | next [-]

GLM 5.3 is very impressive and definitely better, but it also at least 4x the price.

On that note I’ve been subbing in MiMo-2.6-pro when cost is an issue, which is super cheap and also performing really well.

▲pjerem 2 hours ago | parent | prev | next [-]

IDK what happened today but I used GLM-5.3 as usual from Ollama cloud and it was so fast it generated entire documents like instantly.

The reasoning and the result document were done after less than 1 or 2 seconds.

Have Ollama suddenly bought GPU capacity?

▲LeBit 39 minutes ago | parent | prev [-]

There is also GLM 5.3 Flash