From their docs "After using 10M input (cache miss) tokens of MiMo-V2.5-Pro, it is equivalent to consuming 3000M Credits, and you can still enjoy 1100M Credits of MiMo-V2.5". So it's around 12M input credit vs Earlier 60M tokens.
Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%
81–90 of 165 posts
Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%
#82Everyone already said what I wanted to say. That all US companies (OpenAI, Anthropic, Google, MS Copilot) have increased price recently while Chinese companies (Deepseek, Xiaomi) are reducing price. The question is how they are managing to do so? They are supposed to struggle due to chip sanctions. Secondly, why now? The US companies were supposed to subsidize too but now they are unable to keep up. Everyone going to…
As Jensen has been pointing out for almost a year now, these sanctions were ineffective and probably had the opposite effect of the desired goal.
The history is fairly long, but an inflection point could likely be traced to Trump v1 era DOJ enforcement on (among others) Huawei's CFO Meng Wanzhou in 2018. Huawei was hit with the (really big) stick in international transactions: OFAC violation accusations, and it was a seminal moment in the company's internal operations -- they concluded they needed a fully internal supply chain in China, and retooled for it. Meng Wanzhou cases in the US were eventually dismissed, but she was on house arrest in Canada through 2021 or so.
Fast forward to 2024 -- Huawei was culturally and technically ready to build AI accelerators -- one of the externalities of the sanctions was to provide additional benefit to Chinese companies for buying from Huawei; those economics seem to have provided a boost to on-shore development.
Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%
#83Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%
#84Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%
#85Earlier quoted context omitted.
China has a population of 1.4B, US is 349M. 0-8 Beijing time is their off-peak? How is that funny, that's literally how timezones work?
It's funny, in a good way, because their off-peak times match perfectly the werstern peak demand.
Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%
#86Earlier quoted context omitted.
Can't be mistaken for someone like, ugh... Anthropic and OpenAI...
This sort of pressure will force them to though.
Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%
#87as someone from the 3rd world - this is pleasant - even 3rd world countries will have affordable "A.I" access via Chinese models. as someone who now lives & has lived in the west for the majority of their adult life - yeah the US western models r fucked n the crazy valuations of the A.I labs - which also filters down to the economy - since all money instead of being put to productive use is being wasted on this shit.…
I stopped tagging my country as developing and then third world and call it for what it is, a POOR country. I know with increasing certainty that my country will be poor for the rest of my life. I also expect AI to be as available as computers: there are the "have", and there are the "don't have", which is almost always a lifetime condition.
Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%
#88That's deliberate. US AI companies have no chance of recouping even fraction of their valuations. PS: Have not tried this but Deepseev4 Flash (not even Deepseekv4 Pro version) with set to "high" has pretty much Claud Opus 4.7 level of capabilities and is lightening fast and dirty cheap. Hours and hours of conversation barely costs few cents.
A big caveat here is that many US companies (particularly in sensitive industries, like defense) will likely not want to (or not be allowed to) use Chinese models for anything of substance.
Re: Xiaomi MiMo-v2.5 Series API Permanent Price Reduction Up to 99%
#89That's deliberate. US AI companies have no chance of recouping even fraction of their valuations. PS: Have not tried this but Deepseev4 Flash (not even Deepseekv4 Pro version) with set to "high" has pretty much Claud Opus 4.7 level of capabilities and is lightening fast and dirty cheap. Hours and hours of conversation barely costs few cents.
I have been using DeepSeek API within Claude Code. So far it has been legitimately superior to Claude, and Codex that I used before.