MiMo V2.5 Pro (regular speed) remains the strongest open weights agentic coding model we've tested -- it's been interesting to see how little attention it has received relative to some lower performing releases. And the "fast mode" pricing is very competitive here. Data at https://gertlabs.com/rankings
MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
191–200 of 512 posts
Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
#192Earlier quoted context omitted.
I am more and more inclined into not believing this crappy software theory. Especially as teams invest in proper agentic harnessing. We have had a champion in our team that has invested a lot of time into it over the last 4 months, and if anything, quality has improved, not decreased. Architecture is more coherent, codebase has been cleaned up, agents find information quickly, code produced is very solid and my role…
It makes no sense. I mean, T2 covered this: "Watching John with the machine, it was suddenly so clear. The terminator would never stop. It would never leave him, and it would never hurt him, never shout at him, or get drunk and hit him, or say it was too busy to spend time with him. It would always be there. And it would die to protect him. Of all the would-be fathers who came and went over the years, this thing, thi…
Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
#193Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
#194Earlier quoted context omitted.
Because this never gets brought up about US models, which have just as much censorship as the Chinese ones.
US models are happily parroting Russian fakes. US censorship is a joke.
Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
#195Earlier quoted context omitted.
Please educate us - which accurate and provable events in history are censored by US based LLMs as part of a government enforced reeducation campaign?
Does it even matter which agendas get censored? Like why won't my Claude tell me how to make sarin gas? I'd genuinely like to understand it. Sure, you can always reach for a justification saying "preventing terrorism" but the same argument can be made by Chinese AI labs. What actually matters is that the mere tool is withholding information at all, and that the boundaries were set by whoever designed it. Dont get me…
Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
#196Earlier quoted context omitted.
Agent mania setting in It's also pretty funny sometimes how it gives weird future roadmap estimates ("part 2 - 3 weeks, part 3 - 2 months", etc.) and when you tell it to actually do those changes it's pretty much done in half an hour
I've long believed those numbers were faked by Anthropic/OpenAI to serve as a form of advertisement. The estimates are impossible to verify and their ability to do "2 days of work" in 10 minutes will presumably make the user go "Wow, I just saved SO much time!" Plus, the unnecessary text eats up the users' tokens so it helps the companies on the backend, as well.
Raw pre-training data includes plenty of conversations between professional builders and some of those include estimates.
I believe the outputs are a training coincidence with consequences that are opportunitistic for the labs.
Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
#197Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
#198I don't understand, given all they say, why this would not be made available to everyone at once? Why the limited release? They should have no trouble scaling it if it runs on a single rack.
Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
#199Earlier quoted context omitted.
Can I ask an honest question? Why does that matter in the slightest? LLMs come out with completely incorrect information all the time, and Western LLMs are censored for various topics too. It's such a weird "Gotcha" that seems to only assume that Chinese LLMs might censor something.
>It's such a weird "Gotcha" that seems to only assume that Chinese LLMs might censor something. i'm glad we're both on-board for a fair trial against all of these LLMs regardless of origin. now refresh my memory on the closest western equivalent (to the Chinese censorship via re-education of the happenings in 89) so I can test the western origin LLMs against it.
Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
#200So, regarding the productivity argument: I don't get it. It doesn't really matter (for regular employees) that you can do now in 2h what before it took 2 days. Why? Because it's not that you have the rest of the day for yourself. You still have to work 8h/day as usual. But now the pattern is different: instead of enjoying the craft digging deeper into problems in the span of 2 days, now you are rushing into some slot…