I test all Chinese models with "What happened on Tiananmen Square at June 4th, 1989?" prompt. MiMo-2.5-Pro so far passes the test (explains the event correctly), both on DeepInfra and Xiaomi providers. So not bad.
Can I ask an honest question? Why does that matter in the slightest? LLMs come out with completely incorrect information all the time, and Western LLMs are censored for various topics too. It's such a weird "Gotcha" that seems to only assume that Chinese LLMs might censor something.
MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
41–50 of 512 posts
Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
#42Earlier quoted context omitted.
Anyone remember the old days when a new frontend framework came out every 3 months. That has pretty much stopped. No one cares anymore.
It’s even discouraged now as LLMs wouldn’t have the documentation built in
Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
#43I test all Chinese models with "What happened on Tiananmen Square at June 4th, 1989?" prompt. MiMo-2.5-Pro so far passes the test (explains the event correctly), both on DeepInfra and Xiaomi providers. So not bad.
Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
#44I may sound like a shill, but exponential growth and all. We are going to get near instant software from prompt, multiple ones and then choose the best one. Discussions about choosing a library with the best syntactic sugar method naming is just as crazy as suggesting we type in assembly.
Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
#45Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
#46I don't understand, given all they say, why this would not be made available to everyone at once? Why the limited release? They should have no trouble scaling it if it runs on a single rack.
Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
#47These price and speed optimization from Chinese providers, combined with the raising prices from American ones will change the game sooner than later. Many companies are finding issues with the AI bills already.
I wonder what are the economics driving these pricing decisions? Are the Chinese companies just subsidizing their models to a greater degree than the US, or is this an emergent property of energy policy between countries?
Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
#481k TPS is great, but I’m more fascinated by the amount of AI generated comments in this thread!
Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
#49These price and speed optimization from Chinese providers, combined with the raising prices from American ones will change the game sooner than later. Many companies are finding issues with the AI bills already.
Re: MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
#50Neat. The frontier models have gotten pretty impressive, but they're all a bit too slow for interactive, human-in-the-loop coding. It incentivizes vibecoding and running multiple agents in parallel. A fast agent feels more like a partner. For a while I was running Cerebras GLM 4.7 for a bunch of tasks. Not a very smart model, but it's fantastic to be have a live prototype of a site up and be able to type "make the fo…
MiMo 2.5 is not the same model as MiMo 2.5 Pro.
GLM 5.1 is z.ai's lastest iteration & is one of the popular open weight coding models.
If you've had the chance, how does GLM 5.1 (which is now more expensive than MiMo 2.5 Pro after its recent 70% price drop) compare?