Live data from Hacker News

Xiaomi Mimo 2.6 live post-training dashboard

mimo.xiaomi.com

61–70 of 146 posts

Re: Xiaomi Mimo 2.6 live post-training dashboard

#61

For reference, Mimo-v2.5-Pro scored 19% on DeepSWE 1.1. This is looking great. Fable scores 70%, Kimi K3 69%, Astra 74% (all on max effort). https://deepswe.datacurve.ai/blog/deepswe-v1-1

2.6-pro just reached 63.7% by step 10, it's on step 11 right now.

Even flash reached 60.7% by step 12, and it's on step 16 now.

This is so exciting lmao.

Re: Xiaomi Mimo 2.6 live post-training dashboard

#63

Well, if open source AI is dangerous (for OpenAI/Anthropic IPOs?), this is like watching a time bomb.

For my own usage, Luna is cheap enough that I don't care if other models are cheaper. I'm interested if another model is in some way better and not too expensive.

Re: Xiaomi Mimo 2.6 live post-training dashboard

#66

I been using MiMo-V2.5 to do most of my work as software engineer, on a variety of projects I'm working on, and I been VERY happy with ROI. The model is very powerful! Not perfect – I've run in hallucination loops once or twice, but nothing a stop-then-continue wouldn't solve. The cost is unbelievably low, and the quality of intelligence I get is equivalent to when I was working mostly with Anthropic models (late las…

I’ve been very pleased with DS 4.1 flash. Not so much the 4.0 models, but for coding (Rust) it’s been great so far (3 solid days of work).

I’ll give Mimo a try.

Re: Xiaomi Mimo 2.6 live post-training dashboard

#67
post #33

Earlier quoted context omitted.

With other software, devs convince their managers of the importance of using open source stuff in their stack. With AI, it's usually managers choosing what models to use for the devs. The US labs don't need to give a damn how much devs like open source

This isn't about liking open source. This is about the labs just being cool and doing cool shit instead of the opposite which is Anthropic where all they talking about is killing everyone and taking everyone's job.

These labs are still (for the time being) made of people, who reflect their lives onto the work.

The US population is much more pessimistic and doomsday driven these days, whereas the Chinese are more optimistic and future driven.

Re: Xiaomi Mimo 2.6 live post-training dashboard

#68

Well, if open source AI is dangerous (for OpenAI/Anthropic IPOs?), this is like watching a time bomb.

For my own usage, Luna is cheap enough that I don't care if other models are cheaper. I'm interested if another model is in some way better and not too expensive.

Luna is great but makes a lot of mistakes at high and lower in my experience (large rust codebase). I use Luna Max for asynchronous subagent reviews and am very happy with its work, but it’s slow af.
Post reply on HN