Remix.run Logo
joelwallis 9 hours ago

I been using MiMo-V2.5 to do most of my work as software engineer, on a variety of projects I'm working on, and I been VERY happy with ROI. The model is very powerful! Not perfect – I've run in hallucination loops once or twice, but nothing a stop-then-continue wouldn't solve.

The cost is unbelievably low, and the quality of intelligence I get is equivalent to when I was working mostly with Anthropic models (late last year/early this year). I'm fully invested in MiMo and I'm very happy with it.

-- PS: I also check almost daily to see if other models are capable of doing such great work. And they do – DS4F is powerful and DS41 is impressive, GLM 5.3 Flash gets a job done well, etc. – but when I add cost of M-token in the ROI math, Jeez! MiMo is an order of magnitude better.

walrus01 9 hours ago | parent | next [-]

I've found that mimo v2.5 works for very basic things like a python script to do one thing, but it also is very 'dumb' compared to qwen 3.8-flash-next (I think the benchmark scores for terminal and coding specific benches back this up). And definitely not in the same class as like a GLM5.2 or 5.3. It's fast but makes basic mistakes that only get caught later.

girvo 3 hours ago | parent | next [-]

The fact I can run Qwen 3.8 Flash Next locally, forever (on my DGX Spark-alike) is genuinely shocking to me. It’s crazy good for how small it is. Fast, too.

jonsoft an hour ago | parent | next [-]

I made this 3D game in a day on the same setup with Qwen Code as agent: https://games.jonathanpage.com/

And I am not a web developer! It's an extraordinary model.

(Mouse and keyboard required)

walrus01 3 hours ago | parent | prev [-]

Yeah, I'm guessing you have a variant that fits in <128GB with 262k context? I have the unsloth Q8 GGUF of it here in a setup that with full context and ton of extra llama-server "--cache-ram" sits around 200GB RAM usage on a 256GB system, it's probably the best thing I've found for a 256GB class machine. Enough headroom for a rope/yarn extension to 524288 context if I need it.

girvo 2 hours ago | parent [-]

Yep, the engrams are on NVMe (the speed penalty was lower than I expected) and it is quantised to fit.

It’s good enough that I’m considering a second spark, or selling this and buying an M5 Ultra with 256GB for it

7 hours ago | parent | prev [-]
[deleted]
rapind 4 hours ago | parent | prev | next [-]

I’ve been very pleased with DS 4.1 flash. Not so much the 4.0 models, but for coding (Rust) it’s been great so far (3 solid days of work).

I’ll give Mimo a try.

trollbridge 4 hours ago | parent [-]

MiMo is my backup whenever DeepSeek is down, had the price bump, is slow, etc.

UltraSpeed was absolutely awesome. I miss it.

DS 4.1 Flash is amazing. Well worth the extra cost.

flexagoon 6 hours ago | parent | prev | next [-]

How does it compare with DS 4.1 Flash in your experience, if you ignore the cost?

alwinaugustin 6 hours ago | parent | prev | next [-]

I am also using 2.5 and it is giving me solid results. Its available free on Openrouter

jwpapi 8 hours ago | parent | prev | next [-]

May I ask why you ended up there instead of just using the heavy subsidized subscription. I’m actually curious.

eli 6 hours ago | parent [-]

Mimo has subsidized subscriptions too

james2doyle 9 hours ago | parent | prev | next [-]

2.5 Pro or the regular 2.5?

I always found that those Mimo models to be really good at tool calling and following instructions

esafak 8 hours ago | parent | prev | next [-]

How fast is it compared with the other Chinese models?

ricardobeat 7 hours ago | parent [-]

They both are in the 50-100 tok/s range. The Mimo v2.5 Pro Ultraspeed beta could reach 1000 tok/s, hoping they can do something similar for the new model, it was amazing.

electroglyph 6 hours ago | parent | prev | next [-]

[flagged]

NuclearPM 6 hours ago | parent [-]

Real?

electroglyph 3 hours ago | parent [-]

mimo 2.5 has been a big underperformer since shortly after it's release imo. i cancelled my sub after the first month. purposefully using 2.5 right now is just handicapping yourself for no reason.

NuclearPM 2 hours ago | parent [-]

I understand now. You used the wrong word.

yeeeloit 8 hours ago | parent | prev [-]

[flagged]

senordevnyc 8 hours ago | parent | next [-]

Yeah, this Brazilian dude who has been a contributor here on HN longer than your anonymous account is shilling for a Chinese model company. Makes sense.

platinumrad 8 hours ago | parent | prev [-]

Are you accusing them of astroturfing? Why is it strange for someone to say something topical?