Live data from Hacker News

Mistral Medium 3.5

mistral.ai

71–80 of 248 posts

Re: Mistral Medium 3.5

#71
post #37

Earlier quoted context omitted.

Where are the competitive models from Singapore, Japan, Taiwan, Korea, Russia, Canada, India, the UK? From anywhere that isn't China or the US? There are none. Mistral Small 4 is pareto-competitive in its pricing bracket at $0.15/$0.60, at worst it's second to Gemma 4 26B A4B. The above countries have never had a model that is even close to being so. This particular Mistral Medium looks to be uncompetitive at that pr…

> Korea EXAONE from LG AI Research https://huggingface.co/LGAI-EXAONE They had one of the best small models a few months ago and they released a new model just last week. There's also HyperCLOVA X (haven't tested it, but maybe it is also good) https://huggingface.co/naver-hyperclovax > India India has the Sarvam model series, which admittedly are not SotA, but they have pretty good voice capabilities https://huggingf…

I'm familiar with those models. They're nowhere near competitive. Miles away from Mistral or (obviously) Chinese models.

> (haven't tested it, but maybe it is also good)

I have. It is not.

Re: Mistral Medium 3.5

#72
Given what Vibe already did in the previous versions with codestral-v2, that's great news. Keep up the good work ! I don't want to depend on the world's two hungry superpowers.

Re: Mistral Medium 3.5

#73
post #52

I'm not sure what people are on in the comments. It doesn't beat the other models, but it sure competes despite its size. GLM 5.1 is an excellent model, but even at Q4 you're looking at ~400GB. Kimi K2.5 is really good too, and at Q4 quantization you're looking at almost ~600GB. This model? You can run it at Q4 with 70GB of VRAM. This is approaching consumer level territory (you can get a Mac Studio with 128GB of RAM…

I was hoping a lot from it... but this one, is not up to that mark. For example, here is it's comparion with 4.7x smaller model, qwen3.7-27b.

https://chatgpt.com/share/69f239e8-7414-83a8-8fdd-6308906e5f...

Tldr: qwen3.6-27b, a 4.7x smaller model, have similar performance.

Re: Mistral Medium 3.5

#75
post #37

Earlier quoted context omitted.

Where are the competitive models from Singapore, Japan, Taiwan, Korea, Russia, Canada, India, the UK? From anywhere that isn't China or the US? There are none. Mistral Small 4 is pareto-competitive in its pricing bracket at $0.15/$0.60, at worst it's second to Gemma 4 26B A4B. The above countries have never had a model that is even close to being so. This particular Mistral Medium looks to be uncompetitive at that pr…

> Korea EXAONE from LG AI Research https://huggingface.co/LGAI-EXAONE They had one of the best small models a few months ago and they released a new model just last week. There's also HyperCLOVA X (haven't tested it, but maybe it is also good) https://huggingface.co/naver-hyperclovax > India India has the Sarvam model series, which admittedly are not SotA, but they have pretty good voice capabilities https://huggingf…

they should ask unsloth to follow them. For my usecases locally w/128GB, Qwen3.5-Coder-Next is SOTA.

Re: Mistral Medium 3.5

#76
post #32

Earlier quoted context omitted.

Can't agree at all. Productivity gap just 1 year ago was much larger for frontier model vs non-frontier. Let alone 2 years ago.

When I was thinking pre-agentic, I was actually thinking more pre-"coding seen as the main use case for these models".

Coding has always been the main real-world business usecase since day one. There has been no point since the very first public availability of GPT 3.5 in November 2022, that it wasn't.

A lot of us have been agentic coding since almost 2 years ago, mid-2024. I have. The productivity gap of "best vs 2nd vs 3rd best model" was biggest back then and has slowly been shrinking ever since.

Re: Mistral Medium 3.5

#77
post #52

I'm not sure what people are on in the comments. It doesn't beat the other models, but it sure competes despite its size. GLM 5.1 is an excellent model, but even at Q4 you're looking at ~400GB. Kimi K2.5 is really good too, and at Q4 quantization you're looking at almost ~600GB. This model? You can run it at Q4 with 70GB of VRAM. This is approaching consumer level territory (you can get a Mac Studio with 128GB of RAM…

I was hoping a lot from it... but this one, is not up to that mark. For example, here is it's comparion with 4.7x smaller model, qwen3.7-27b. https://chatgpt.com/share/69f239e8-7414-83a8-8fdd-6308906e5f... Tldr: qwen3.6-27b, a 4.7x smaller model, have similar performance.

That's a chatgpt summary. Actual usage would a better test.

Re: Mistral Medium 3.5

#78
post #51
post #37

Earlier quoted context omitted.

Where are the competitive models from Singapore, Japan, Taiwan, Korea, Russia, Canada, India, the UK? From anywhere that isn't China or the US? There are none. Mistral Small 4 is pareto-competitive in its pricing bracket at $0.15/$0.60, at worst it's second to Gemma 4 26B A4B. The above countries have never had a model that is even close to being so. This particular Mistral Medium looks to be uncompetitive at that pr…

DeepMind, which is headquartered in London, probably had a significant role in the development of the Gemini and Gemma models. Yes, it might be a problem that the UK allows companies like this to be bought up by foreign countries.

Without Google’s funding its not obvious i DeepMind would have went anywhere.

Unless the moved to US for funding while keeping a back office in the UK.

It’s strange to expect anything significant to come out from Europe when VCs there are either very risk averse and/or don’t have enough cash to begin with. It’s not like government or EU funding can replace that since its almost always wasted or missdirected

Re: Mistral Medium 3.5

#79
post #52

I'm not sure what people are on in the comments. It doesn't beat the other models, but it sure competes despite its size. GLM 5.1 is an excellent model, but even at Q4 you're looking at ~400GB. Kimi K2.5 is really good too, and at Q4 quantization you're looking at almost ~600GB. This model? You can run it at Q4 with 70GB of VRAM. This is approaching consumer level territory (you can get a Mac Studio with 128GB of RAM…

I was hoping a lot from it... but this one, is not up to that mark. For example, here is it's comparion with 4.7x smaller model, qwen3.7-27b. https://chatgpt.com/share/69f239e8-7414-83a8-8fdd-6308906e5f... Tldr: qwen3.6-27b, a 4.7x smaller model, have similar performance.

To be fair MoE from Qwen itself had the same "problem". 3.5 122B MoE was same or worse than 3.5 27B. Yet to see 122B 3.6.

UPD. NVM, Mistral Medium 3.5 is dense. So yes, it is worse in every way.

Re: Mistral Medium 3.5

#80
post #71

Earlier quoted context omitted.

> Korea EXAONE from LG AI Research https://huggingface.co/LGAI-EXAONE They had one of the best small models a few months ago and they released a new model just last week. There's also HyperCLOVA X (haven't tested it, but maybe it is also good) https://huggingface.co/naver-hyperclovax > India India has the Sarvam model series, which admittedly are not SotA, but they have pretty good voice capabilities https://huggingf…

I'm familiar with those models. They're nowhere near competitive. Miles away from Mistral or (obviously) Chinese models. > (haven't tested it, but maybe it is also good) I have. It is not.

You mentioned "pareto-competitive", and EXAONE certainly was that. The statement that the "above countries have never had a model that is even close to being so" is simply too broad.
Post reply on HN