logoalt Hacker News

joelwallisyesterday at 8:54 PM14 repliesview on HN

I been using MiMo-V2.5 to do most of my work as software engineer, on a variety of projects I'm working on, and I been VERY happy with ROI. The model is very powerful! Not perfect – I've run in hallucination loops once or twice, but nothing a stop-then-continue wouldn't solve.

The cost is unbelievably low, and the quality of intelligence I get is equivalent to when I was working mostly with Anthropic models (late last year/early this year). I'm fully invested in MiMo and I'm very happy with it.

-- PS: I also check almost daily to see if other models are capable of doing such great work. And they do – DS4F is powerful and DS41 is impressive, GLM 5.3 Flash gets a job done well, etc. – but when I add cost of M-token in the ROI math, Jeez! MiMo is an order of magnitude better.


Replies

ehsankiatoday at 9:02 AM

> late last year/early this year

That's an eternity when it comes to coding models.

In my personal experience, we've had almost a step change every ~3 months this year, at least for bigger one-shot tasks. For example looking at Gemini Flash 3.0 vs 3.5 vs 3.8, it went 5% -> 30% -> 75% on DeepSWE, all since the start of the year.

miyurutoday at 5:59 AM

Same here. It’s the first AI provider I actually gave money to, since they offered the model for free with a Mimo code for the first month or so, and it was great.

These days, there are more intelligent models like DS4.1, but Mimo is very obedient, so I plan things with another model and give the implementation to Mimo.

alwinaugustinyesterday at 11:21 PM

I am also using 2.5 and it is giving me solid results. Its available free on Openrouter

rapindtoday at 1:25 AM

I’ve been very pleased with DS 4.1 flash. Not so much the 4.0 models, but for coding (Rust) it’s been great so far (3 solid days of work).

I’ll give Mimo a try.

show 1 reply
baxtrtoday at 6:31 AM

Could you elaborate on how you check daily? Do you swap models for certain tasks?

walrus01yesterday at 9:15 PM

I've found that mimo v2.5 works for very basic things like a python script to do one thing, but it also is very 'dumb' compared to qwen 3.8-flash-next (I think the benchmark scores for terminal and coding specific benches back this up). And definitely not in the same class as like a GLM5.2 or 5.3. It's fast but makes basic mistakes that only get caught later.

show 2 replies
jwpapiyesterday at 9:37 PM

May I ask why you ended up there instead of just using the heavy subsidized subscription. I’m actually curious.

show 1 reply
flexagoonyesterday at 11:45 PM

How does it compare with DS 4.1 Flash in your experience, if you ignore the cost?

james2doyleyesterday at 8:58 PM

2.5 Pro or the regular 2.5?

I always found that those Mimo models to be really good at tool calling and following instructions

ignoramoustoday at 6:46 AM

> GLM 5.3 Flash gets a job done well, etc. – but when I add cost of M-token in the ROI math, Jeez! MiMo is an order of magnitude better.

API may be expensive, but I do 900m tokens (95% cached, ~0.4% output) on Z.ai's $18/mo coding plan with GLM 5.3 Flash.

esafakyesterday at 9:55 PM

How fast is it compared with the other Chinese models?

show 1 reply
wangxili1997today at 8:05 AM

[flagged]

electroglyphyesterday at 11:50 PM

[flagged]

show 1 reply
yeeeloityesterday at 9:38 PM

[flagged]

show 2 replies