Live data from Hacker News

MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second

mimo.xiaomi.com

221–230 of 512 posts

Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second

#221
An exercise for the near future:

Albert has a chalet in swiss alps and an uncles' fortune, burning tokens at 11 kHz.

Joe has a rental capsule and a UBI, burning equally priced tokens at 23kHz.

Who's the first to solve the problem of maniacs in power?

Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second

#223

Earlier quoted context omitted.

Agent mania setting in It's also pretty funny sometimes how it gives weird future roadmap estimates ("part 2 - 3 weeks, part 3 - 2 months", etc.) and when you tell it to actually do those changes it's pretty much done in half an hour

I've long believed those numbers were faked by Anthropic/OpenAI to serve as a form of advertisement. The estimates are impossible to verify and their ability to do "2 days of work" in 10 minutes will presumably make the user go "Wow, I just saved SO much time!" Plus, the unnecessary text eats up the users' tokens so it helps the companies on the backend, as well.

All models do it. It's their training. They didn't have "a person does this in a week but an LLM could in a minute" in their training yet. They also don't have the concept of elapsed time unless you ask them how long something has taken.

Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second

#224
post #58

Fast AI seems genuinely exciting and somewhat unsettling to me. Right now Claude is faster than me on some tasks but we’re at least close. I have a prompt to clean up a PR that’s been running for 1h now and I expect it to take another few. It’s hard to imagine how the workflow would look like if it was near-instant. On the one hand, it might be easier to focus. Some prompts take so long that I start to multitask and…

Have you tried Gemini 3.5 Flash? It's quite fast. Amazing how fast it finishes tasks. Much faster than Claude.

Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second

#225
post #44

Earlier quoted context omitted.

Sounds like exponential growth of crappy software. I'm not saying that before we didn't have mass produced crap in SE, but now it will turn into explosive overflow.

"exponential growth of crappy X" applies to every industry that went from being an artisanal craft to being mass produced with little or no human input. and we live much better lives than we did before the industrial revolution.

I think some industries have notably high quality output. Automobiles, aerospace for example.

Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second

#226
post #220
post #58

Fast AI seems genuinely exciting and somewhat unsettling to me. Right now Claude is faster than me on some tasks but we’re at least close. I have a prompt to clean up a PR that’s been running for 1h now and I expect it to take another few. It’s hard to imagine how the workflow would look like if it was near-instant. On the one hand, it might be easier to focus. Some prompts take so long that I start to multitask and…

> Right now Claude is faster than me on some tasks but we’re at least close. I dont doubt it, but I don't think you can spawn 10 copies of yourself working simultaneously.

No, but nor can you keep track of what 10 agents are doing simultaneously. Hence the multitasking regret.

Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second

#227
post #44

Earlier quoted context omitted.

Sounds like exponential growth of crappy software. I'm not saying that before we didn't have mass produced crap in SE, but now it will turn into explosive overflow.

Crap is fine if it gets the job done. I think software as an industry will change to more ephemeral construction.

Paper plates of software development.

Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second

#228
post #76

Earlier quoted context omitted.

Another problem is that US models are all closed source, and if you're a large corporate you may not want your org to be held hostage by OpenAI / Anthropic. I genuinely don't understand what moat these US model labs have. If they're saying recursive self improvement is just around the corner and Chinese labs are only slightly behind the leading US models, what moat does the US labs have? Are the US models going to re…

> you may not want your org to be held hostage by OpenAI / Anthropic Or Google. I'm working with multiple customers right now that are very pissed at Google for deprecating Gemini 2.5 Flash, canning the GA release of 3.0 Flash and now have to decide whether to bite the bullet of the 5x price increase for 3.5 Flash or switching providers. Quite a few of them will likely fully pivot to open models.

I'd be curious if any of your customers have tried 3.1 Flash Lite. It's cheaper than 2.5 Flash, and in my experience with the free tier, quite an upgrade in terms of quality of response. My suspicion is that Google is killing off the old models because they aren't a good value for the customer or for themselves.

Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second

#229

MiMo V2.5 Pro (regular speed) remains the strongest open weights agentic coding model we've tested -- it's been interesting to see how little attention it has received relative to some lower performing releases. And the "fast mode" pricing is very competitive here. Data at https://gertlabs.com/rankings

why is deepseek v4 pro a lot lower than flash? where is mimo 2.5?

[deleted]

Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second

#230
post #76

Earlier quoted context omitted.

Another problem is that US models are all closed source, and if you're a large corporate you may not want your org to be held hostage by OpenAI / Anthropic. I genuinely don't understand what moat these US model labs have. If they're saying recursive self improvement is just around the corner and Chinese labs are only slightly behind the leading US models, what moat does the US labs have? Are the US models going to re…

I think they are racing because the first ASI will 'win', preventing others, of course we won't be able to bake the right goals into it though.

i dont think its going to automatically prevent others. super claude might understand why diversity is important. if were talking sci fi scenarios the most likely one is probably overwatch (multiple independent ais with gray ethics and complicated relationships) more than skynet.
Post reply on HN