These benchmark numbers cannot be real for a 7b model
Xiaomi MiMo Reasoning Model
41–50 of 203 posts
Re: Xiaomi MiMo Reasoning Model
#42Earlier quoted context omitted.
But who will keep them updated and what incentive they would have? That's I can't imagine. Bit vague.
Who keeps open source projects maintained and what incentive do they have?
Re: Xiaomi MiMo Reasoning Model
#43[flagged]
Source (Chinese): https://finance.sina.cn/tech/2020-11-26/detail-iiznctke33979...
Re: Xiaomi MiMo Reasoning Model
#44Earlier quoted context omitted.
Last time I did that I was also impressed, for a start. Problem was that of a top ten book recommendations only the first 3 existed and the rest was a casually blended hallucination delivered in perfect English without skipping a beat. "You like magic? Try reading the Harlew Porthouse series by JRR Marrow, following the orphan magicians adventures in Hogwesteros" And the further towards the context limit it goes the…
LLMs are not search engines…
Re: Xiaomi MiMo Reasoning Model
#45Earlier quoted context omitted.
The smaller models have been creeping upward. They don't make headlines because they aren't leapfrogging the mainline models from the big companies, but they are all very capable. I loaded up a random 12B model on ollama the other day and couldn't believe how good it competent it seemed and how fast it was given the machine I was on. A year or so ago, that would have not been the case.
What model? I have been using api's mostly since ollama was too slow for me.
[0]: https://huggingface.co/mlabonne/gemma-3-12b-it-abliterated
Re: Xiaomi MiMo Reasoning Model
#46[flagged]
Re: Xiaomi MiMo Reasoning Model
#47Re: Xiaomi MiMo Reasoning Model
#48Re: Xiaomi MiMo Reasoning Model
#49Re: Xiaomi MiMo Reasoning Model
#50Waiting for GGUF or MLX models. Probably within few hours will be released.