Live data from Hacker News

Xiaomi Mimo 2.6 live post-training dashboard

mimo.xiaomi.com

71–80 of 150 posts

Re: Xiaomi Mimo 2.6 live post-training dashboard

#71
post #66

I been using MiMo-V2.5 to do most of my work as software engineer, on a variety of projects I'm working on, and I been VERY happy with ROI. The model is very powerful! Not perfect – I've run in hallucination loops once or twice, but nothing a stop-then-continue wouldn't solve. The cost is unbelievably low, and the quality of intelligence I get is equivalent to when I was working mostly with Anthropic models (late las…

I’ve been very pleased with DS 4.1 flash. Not so much the 4.0 models, but for coding (Rust) it’s been great so far (3 solid days of work). I’ll give Mimo a try.

MiMo is my backup whenever DeepSeek is down, had the price bump, is slow, etc.

UltraSpeed was absolutely awesome. I miss it.

DS 4.1 Flash is amazing. Well worth the extra cost.

Re: Xiaomi Mimo 2.6 live post-training dashboard

#72
I absolutely love that someone is doing this! Why isn’t IBM for Granite or Google for Gemini?

If you are going to develop a near frontier model, and you don’t think you have special sauce up your sleeve, why not making training runs and RL environment scores etc. visible to the world?

I’m genuinely learning quite a bit just from the dashboard

Re: Xiaomi Mimo 2.6 live post-training dashboard

#73

I been using MiMo-V2.5 to do most of my work as software engineer, on a variety of projects I'm working on, and I been VERY happy with ROI. The model is very powerful! Not perfect – I've run in hallucination loops once or twice, but nothing a stop-then-continue wouldn't solve. The cost is unbelievably low, and the quality of intelligence I get is equivalent to when I was working mostly with Anthropic models (late las…

I've found that mimo v2.5 works for very basic things like a python script to do one thing, but it also is very 'dumb' compared to qwen 3.8-flash-next (I think the benchmark scores for terminal and coding specific benches back this up). And definitely not in the same class as like a GLM5.2 or 5.3. It's fast but makes basic mistakes that only get caught later.

The fact I can run Qwen 3.8 Flash Next locally, forever (on my DGX Spark-alike) is genuinely shocking to me. It’s crazy good for how small it is. Fast, too.

Re: Xiaomi Mimo 2.6 live post-training dashboard

#74

You'd think they would make it less obvious that they are running their whole operation with Claude

If you're thinking of the UI style, definitely not Claude. It is incapable of writing a clear sentence like "what each step's samples are made of", would have used all-caps for everything, more padding and gradients.

I hope this is /s because it’s very easy to get Claude to write sensibly. That’s why AI slop writing is so annoying because it’s so easy to avoid with any amount of effort at all.

Re: Xiaomi Mimo 2.6 live post-training dashboard

#78

Very curious that everyone here (so far) seems to assume this dashboard presents real data.

Haha yeah pretty wild how easily you can see the data is fake by the repeating numbers (refresh the page the progress goes back in time constantly) + watch for restarts. They say they happen but 0 data correlates the log messages. Just a replay of old data or being fed by an llm so they convince people they are open

Re: Xiaomi Mimo 2.6 live post-training dashboard

#79

The Chinese labs are just making fun of the US labs at this point. Where is the cool shit from the US labs?

You mean all of the frontier models that the Chinese distillation clones are copying? Yeah kinda cool imo. If a dashboard showing training for a model that doesn't even come close to anything us labs have released in 6 months is "cool", then you're a loser

Re: Xiaomi Mimo 2.6 live post-training dashboard

#80
post #72

I absolutely love that someone is doing this! Why isn’t IBM for Granite or Google for Gemini? If you are going to develop a near frontier model, and you don’t think you have special sauce up your sleeve, why not making training runs and RL environment scores etc. visible to the world? I’m genuinely learning quite a bit just from the dashboard

They think they have the special sauce. Even if they do, what would they get in return for doing that?
Post reply on HN