Live data from Hacker News

Magistral — the first reasoning model by Mistral AI

mistral.ai

31–40 of 444 posts

Re: Magistral — the first reasoning model by Mistral AI

#31
post #19

The featured accuracy benchmarks exclude every model that matter except DeepSeek, which is quite telling about this new model's performance. This makes it yet another example of European companies building great products but fumbling marketing. Mistral's edge is speed. It's a real pleasure to use because it answers in ~1s what takes other models 5-8s, which makes for a much better experience. But instead of focusing…

That is reasonable though. Comparing the product of a small company with little resources with giants like Google and OpenAI in a field where most advances are due to more and more expensive models is nonsense.

The point I was trying to express is that Mistral is arguably far superior to the giants if you care about speed! So I wished they communicated this more clearly.

Re: Magistral — the first reasoning model by Mistral AI

#32
post #11

Benchmarks suggest this model loses to Deepseek-R1 in every one-shot comparison. Considering they were likely not even pitting it against the newer R1 version (no mention of that in the article) and at more than double the cost, this looks like the best AI company in the EU is struggling to keep up with the state-of-the-art.

[deleted]

Re: Magistral — the first reasoning model by Mistral AI

#33

I made some GGUFs for those interested in running them at https://huggingface.co/unsloth/Magistral-Small-2506-GGUF ollama run hf.co/unsloth/Magistral-Small-2506-GGUF:UD-Q4_K_XL or ./llama.cpp/llama-cli -hf unsloth/Magistral-Small-2506-GGUF:UD-Q4_K_XL --jinja --temp 0.7 --top-k -1 --top-p 0.95 -ngl 99 Please use --jinja for llama.cpp and use temperature = 0.7, top-p 0.95! Also best to increase Ollama's context length…

But this is just the SFT - "distilled" model, not the one optimized with RL, right?

Re: Magistral — the first reasoning model by Mistral AI

#34
post #11

Benchmarks suggest this model loses to Deepseek-R1 in every one-shot comparison. Considering they were likely not even pitting it against the newer R1 version (no mention of that in the article) and at more than double the cost, this looks like the best AI company in the EU is struggling to keep up with the state-of-the-art.

Thought so too. I don't know how it could be different though. They are competing against behemoths like OpenAI or Google, but have only 200 people. Even Anthropic has over 1000 people. DeepSeek has less than 200 people so the comparison seems fair.

any claim from the deepseek folks should be considered with wide margins of error.

Re: Magistral — the first reasoning model by Mistral AI

#35
post #11

Benchmarks suggest this model loses to Deepseek-R1 in every one-shot comparison. Considering they were likely not even pitting it against the newer R1 version (no mention of that in the article) and at more than double the cost, this looks like the best AI company in the EU is struggling to keep up with the state-of-the-art.

If you look at Mistral investors[0], you will quickly understand that Mistral is far from being European. My understanding is it is mainly owned by US companies with a few other companies from EU and other places in the world.

[0] https://tracxn.com/d/companies/mistral-ai/__SLZq7rzxLYqqA97j... (edited for typo)

Re: Magistral — the first reasoning model by Mistral AI

#36
post #11

Benchmarks suggest this model loses to Deepseek-R1 in every one-shot comparison. Considering they were likely not even pitting it against the newer R1 version (no mention of that in the article) and at more than double the cost, this looks like the best AI company in the EU is struggling to keep up with the state-of-the-art.

"EU is leading in regulation", they say. I don't know what they are thinking.

Sorry, this is just getting old...

Its a trite talking point and not the reason why there are so few consumer-AI companies in Europe.

Re: Magistral — the first reasoning model by Mistral AI

#37
post #24

Earlier quoted context omitted.

"EU is leading in regulation", they say. I don't know what they are thinking.

It is fairly common to struggle to understand why different cultures think the way they do.

I live in Europe.

Re: Magistral — the first reasoning model by Mistral AI

#38
post #21
post #11

Benchmarks suggest this model loses to Deepseek-R1 in every one-shot comparison. Considering they were likely not even pitting it against the newer R1 version (no mention of that in the article) and at more than double the cost, this looks like the best AI company in the EU is struggling to keep up with the state-of-the-art.

Europe isn't going to catch up in tech as long as its market is open to US tech giants. Tech doesn't have marginal costs, so you want to have one of it in one place and sell it everywhere and when the infra and talent is already in US, EU tech is destined to do niche products. UK has a bit of it, France has some and that's it. The only viable alternatives are countries who have issues with US and that is China and Ru…

If you close off the market to US tech giants, maybe they'll have some amount of market dominance at home, but I would doubt that would mean they've "caught up" tech wise. There would be no incentive to compete. American EV manufacturing is pretty far behind Chinese EV manufacturing, protectionism didn't help make a competitive car, it just protected the home market while slowly ceding international market after international market

Re: Magistral — the first reasoning model by Mistral AI

#39
I don't understand why the benchmark selections are so scattered and limited. It only compares Magistral Medium with Deepseek V3, R1, and the other close weighted Mistral Medium 3. Why did they leave off Magistral Small entirely, alongside comparisons with Alibaba Qwen or the mini versions of o3 and o4?
Post reply on HN