MiniMax M2.5 is trained by Claude Opus 4.6?
1–10 of 15 posts
Re: MiniMax M2.5 is trained by Claude Opus 4.6?
#2Re: MiniMax M2.5 is trained by Claude Opus 4.6?
#3Re: MiniMax M2.5 is trained by Claude Opus 4.6?
#4They all are trained by each other. Claude says it's DeepSeek if you ask it in Mandarin.
Re: MiniMax M2.5 is trained by Claude Opus 4.6?
#5Name training is always shallow, Claude itself would claim it's GPT-3, GPT-4, or Reddit (heh) when confused. It's just dataset contamination, because the web is full of slop. Never trust self-reported names.
Re: MiniMax M2.5 is trained by Claude Opus 4.6?
#6Re: MiniMax M2.5 is trained by Claude Opus 4.6?
#7This has been a common issue with the Chinese open weight models. It appears most or all have been trained via distillation on OpenAI and Anthropic models.
Re: MiniMax M2.5 is trained by Claude Opus 4.6?
#8Re: MiniMax M2.5 is trained by Claude Opus 4.6?
#9You get an open model which is a 95% of Opus 4.6 quality and 80% cheaper in most inference providers and also can run on your own hardware
Also they did the hard parts of:
* crawling the content
* running the fine tuning (or training)
Better than 1 or 2 companies taking control of the whole AI economy
Re: MiniMax M2.5 is trained by Claude Opus 4.6?
#10This has been a common issue with the Chinese open weight models. It appears most or all have been trained via distillation on OpenAI and Anthropic models.
They most likely weren't, despite very dubious claims of Amodei and Altman and a certain twitter influencer running a pretty naive writing benchmark ("slop test") that is wrong in a very obvious manner. The only unambiguous cases of distillation were Gemini 2.0 experimentals being trained on Claude outputs, and GLM-4.7 being trained on Gemini 3.0 Pro. The rest are pretty different from each other.