Live data from Hacker News

Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%

platform.xiaomimimo.com

121–130 of 165 posts

Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%

#121
post #98
post #91

Hot take: The reason this is happening is because the market for Chinese AI models hosted by Chinese companies is struggling. Even the market for Chinese AI models hosted by western companies is soft: During the week of May 18, OpenRouter processed 3.4T DeepSeek v3 Flash tokens (their most popular model). Google has announced that Gemini is processing 746T per week; Claude is probably processing more. And the Chinese…

> OpenRouter processed 3.4T DeepSeek v3 Flash > Gemini is processing 746T per week I read this totally differently. A startup nobody really knows is doing half a percent of Google on a commodity task?!? Google, which puts Gemini on billions of devices by default, without the user asking? Google, which is distributing Gemini to users who are unaware they are even using it? Versus a startup that does not even have a lo…

Agreed.

Not to mention, week on week more and more tokens are being processed via OpenRouter. [0]. The number keeps going up, with no end in sight in my opinion, if the China models continue offering cheaper inference, whilst tailing behind not too far, the line will keep going up.

[0] - https://openrouter.ai/rankings

OpenRouter is not the only "router" type AI company. More fixed providers like OpenCode and commandcode are offering subscription services on open/china models, likely consuming billions of tokens each. Who know how many tokens are being process directly against Deekseek and Kimi's APIs.

Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%

#122

How realistic is this: Chinese models incidentally slurps up some terms that lead them to finding unflattering words that you wrote about the CCP in a random journal entry, or maybe a social media csv export. You go to China one day and are denied entry due to what you said. Realistic or no? (yes i know the us is getting bad in re. to what you write online as well) Models hosted in China are a siren call that I don't…

This statement makes no sense, because you literally said the "US is getting bad". We already gave up all of our data, if you wrote something about the CCP you should already expect they know about it.

Besides that, the us govt already has all your data and yet people are criticising it all around, in the open. They can, without repercussions, because the us is a free country.

Chinese people can’t really do the same.

Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%

#123

Insane. 3 points behind opus on the artificialanalysis index. Mimo cost ~$400 at the old price, so about $40 today. Opus cost ~$5000 That's over 100x cheaper, and just 3 points behind. I can't wait to experiment with an llm consortium of 100 deepseek and mimo models. Crazy times. Shut up and take my m̶o̶n̶e̶y̶ data! Edit: Gemini on google search told me I could write strikethrough text on hn using . Mimo told me it w…

I had a subscription before the price was cut down; the model kept randomly looping the with same character (burning 30% of the budget in one shot), and the overall performance for agentic purposes is, simply put, terrible. It finds non-existing bugs and randomly removes chunks of code to fix them, then even presents it as an "extra fix". Maybe it's a good generalistic model; I haven't tested it in that regard.

MiniMax (currently 2.7) which is a ~270B model tuned exclusively for agentic purposes, performs so MUCH better; it's more reliable and cheaper. Both are still far away from Opus 4.7 that I'm using at work. IMO benchmarks are just a very rough estimation; everyone cheats as much as they can get away with. Test the model yourself; do not make any assumptions based on the benchmarks.

I would love to see specialized, cheaper, bleeding-edge models like MiniMax for other non-agentic purposes as well. Why pay $1 for a general model when, for example, you can pay $0.1 for a content-moderator model that you actually need?

Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%

#124
post #103

Earlier quoted context omitted.

Still no.

Why?

USA is as censored as what we believe about China. You will get cancelled by Americans if you use "communist" stuff. That is why hardly any Chinese EVs in USA. Because it is communists stuff. The odd thing is iphone is made in China. So it is more of selective enforcement when convenience. Chinese AI even self host means your will influence by communism. You want McCarthy era back again?

Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%

#125
post #25

Since the 3rd party providers on openrouter have all converged on much higher prices in serving these models (both mimo and dsv4), there's obviously a question on how/why are they lowering the prices so much. It's possible they've finally integrated cheap(er) chinese chips. It's also possible they're just subsidising inference for real-world usage data. Interesting either way.

> how/why are they lowering the prices so much Like I responded to someone else: - Cheap electricity - Cheap, domestically produced GPUs - Efficiency research. (a lot of it from Deepseek's research) Also, the Chinese government wants the AI to be as accessible as EVs so everyone will use it.

Also if this is on the path of anything the Chinese do in the physical goods world, inference will be rockbottom cheap in a few years because they'll invest in the hell out of energy, GPUs, research, etc. The same thing they did with EVs.

Only artificial barriers will keep people using some of the frontier stuff in a couple of years. No costs will justify.

Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%

#126

Insane. 3 points behind opus on the artificialanalysis index. Mimo cost ~$400 at the old price, so about $40 today. Opus cost ~$5000 That's over 100x cheaper, and just 3 points behind. I can't wait to experiment with an llm consortium of 100 deepseek and mimo models. Crazy times. Shut up and take my m̶o̶n̶e̶y̶ data! Edit: Gemini on google search told me I could write strikethrough text on hn using . Mimo told me it w…

So I tried the $16/mo token plan. Burned through 31% of monthly budget in one 1-2h session of a small C project refactoring, saw some not great behavior (hey subagent, read me back these 6 files exactly - which probably burned a lot of output tokens) and will cancel, obviously.

This is waaaaay more constrained than even Claude Pro plan, let alone Deepseek V4 or Kimi K2.6 pricing.

Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%

#127
post #104

Insane. 3 points behind opus on the artificialanalysis index. Mimo cost ~$400 at the old price, so about $40 today. Opus cost ~$5000 That's over 100x cheaper, and just 3 points behind. I can't wait to experiment with an llm consortium of 100 deepseek and mimo models. Crazy times. Shut up and take my m̶o̶n̶e̶y̶ data! Edit: Gemini on google search told me I could write strikethrough text on hn using . Mimo told me it w…

benchmarks we deserve: google search quick ai answers vs full llm model :)

search answers use Flash 3.5

Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%

#128

Insane. 3 points behind opus on the artificialanalysis index. Mimo cost ~$400 at the old price, so about $40 today. Opus cost ~$5000 That's over 100x cheaper, and just 3 points behind. I can't wait to experiment with an llm consortium of 100 deepseek and mimo models. Crazy times. Shut up and take my m̶o̶n̶e̶y̶ data! Edit: Gemini on google search told me I could write strikethrough text on hn using . Mimo told me it w…

I had a subscription before the price was cut down; the model kept randomly looping the with same character (burning 30% of the budget in one shot), and the overall performance for agentic purposes is, simply put, terrible. It finds non-existing bugs and randomly removes chunks of code to fix them, then even presents it as an "extra fix". Maybe it's a good generalistic model; I haven't tested it in that regard. MiniM…

Funny, I had the opposite experience with MiniMax and Mimo when using OpenCode. MiniMax got stuck with looping through broken tool calls all the time and MiMo just powered through things and for the most part just worked.

Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%

#129
post #91

Hot take: The reason this is happening is because the market for Chinese AI models hosted by Chinese companies is struggling. Even the market for Chinese AI models hosted by western companies is soft: During the week of May 18, OpenRouter processed 3.4T DeepSeek v3 Flash tokens (their most popular model). Google has announced that Gemini is processing 746T per week; Claude is probably processing more. And the Chinese…

comparing deepseek usage on openrouter to google usage in total is not statistically correct

you could equally say, in the last complete week openrouter processed more deepseek tokens than any other provider including google

that also would not tell you much about how many tokens are used on deepseek

Post reply on HN