Live data from Hacker News

Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%

platform.xiaomimimo.com

91–100 of 165 posts

Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%

#91
Hot take: The reason this is happening is because the market for Chinese AI models hosted by Chinese companies is struggling. Even the market for Chinese AI models hosted by western companies is soft: During the week of May 18, OpenRouter processed 3.4T DeepSeek v3 Flash tokens (their most popular model). Google has announced that Gemini is processing 746T per week; Claude is probably processing more. And the Chinese models were already staggeringly cheap, far cheaper than most Gemini, Claude, or GPT models, before this recent array of pricing changes.

Broadly: No one is using the Chinese AI models. Everyone, globally, everywhere, including in China, is using the models from OpenAI, Anthropic, and Google. The models from the Big Three western labs represent >80% of all tokens processed and likely >95% of all revenue.

Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%

#92
post #59

Earlier quoted context omitted.

The state of the art models (mostly GPT 5.5, but also Gemini and Claude) are better so they cost more. Qwen 3.7 Max is their only direct competition and it is not any cheaper.

Are they? I have been using DeepSeek, and I am finding it better than Claude or Codex, to be honest. I don't see myself going back.

I love ds4, us models are better imo, but like 5% not 500% better, so the valuation doesn't really make sense

that being said, deepseek v4 needs to be on amazon bedrock to actually be feasible in the US Enterprise market and start driving other provider prices down

Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%

#93
post #85
post #54

Earlier quoted context omitted.

It's funny, in a good way, because their off-peak times match perfectly the werstern peak demand.

Can folks in China run US-based models? Seems like they should take advantage of this overlap in peak timing.

Yes, use VPN; they are the main clients

Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%

#95
post #88
post #29

That's deliberate. US AI companies have no chance of recouping even fraction of their valuations. PS: Have not tried this but Deepseev4 Flash (not even Deepseekv4 Pro version) with set to "high" has pretty much Claud Opus 4.7 level of capabilities and is lightening fast and dirty cheap. Hours and hours of conversation barely costs few cents.

> US AI companies have no chance of recouping even fraction of their valuations. A big caveat here is that many US companies (particularly in sensitive industries, like defense) will likely not want to (or not be allowed to) use Chinese models for anything of substance.

What about self host Chinese models?

Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%

#96
post #5

I worked part time with MiMo 2.5-pro over the last month, and barely managed to use 500 Million of the 700 Million tokens I had allocated. My plan was just upgraded to 38 BILLION tokens per month. That's at least 10X the tokens I've used in my entire agentic development so far. I should probably downgrade my plan, but we'll see. :)

Token allocation/cost aside, how was the quality of the model? Any comparison with any other model you've used?

For example, I've heard DeepSeek v4 Pro is comparable to Sonnet 4.7, so I just bought some credits to try it out.

Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%

#97
Anything to destroy US tech companies is welcome.

They aren't aiming companies but users which many have no common sense and grant these agentic AI access to everything.

All the restrictions the US imposed to CH, will be reverted back and it will be even worse, because now the data is not reaching the US gov ( we all know they have access to US big techs data ) but CH.

I really hope this goes viral and breaks Nvidia/OpenAI.

Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%

#98
post #91

Hot take: The reason this is happening is because the market for Chinese AI models hosted by Chinese companies is struggling. Even the market for Chinese AI models hosted by western companies is soft: During the week of May 18, OpenRouter processed 3.4T DeepSeek v3 Flash tokens (their most popular model). Google has announced that Gemini is processing 746T per week; Claude is probably processing more. And the Chinese…

> OpenRouter processed 3.4T DeepSeek v3 Flash

> Gemini is processing 746T per week

I read this totally differently. A startup nobody really knows is doing half a percent of Google on a commodity task?!? Google, which puts Gemini on billions of devices by default, without the user asking? Google, which is distributing Gemini to users who are unaware they are even using it?

Versus a startup that does not even have a login button on its homepage?

This is astonishing.

Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%

#99

Since the 3rd party providers on openrouter have all converged on much higher prices in serving these models (both mimo and dsv4), there's obviously a question on how/why are they lowering the prices so much. It's possible they've finally integrated cheap(er) chinese chips. It's also possible they're just subsidising inference for real-world usage data. Interesting either way.

> there's obviously a question on how/why are they lowering the prices so much. Same reason they release some of the models for free: They are trying to capture market share.

LLM providers can't "capture" anything. People loved Claude Code because it was cheap and good. Not cheap anymore? People switching to Codex, DS4 etc.

Their only moat is maybe being SOTA but that only lasts so long before everyone else catches up.

Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%

#100
post #85

Earlier quoted context omitted.

Can folks in China run US-based models? Seems like they should take advantage of this overlap in peak timing.

Yes, use VPN; they are the main clients

Why do they use a VPN?
Post reply on HN