Europe's bright star has been quiet for a while, great to see them back and good to see them come back to Open Source light with Apache 2.0 licenses - they're too far from the SOTA pack that exclusive/proprietary models would work in their favor. Mistral had the best small models on consumer GPUs for a while, hopefully Ministral 14B lives up to their benchmarks.
All thanks to the US VCs that acutally have money to fund Mistral's entire business. Had they gone to the EU, Mistral would have gotten a miniscule grant from the EU to train their AI models.
Mistral 3 family of models released
51–60 of 243 posts
Re: Mistral 3 family of models released
#52Went to actually use it, got a message saying that I missed a payment 8 months previously and thus wasn't allowed to use Pro despite having paid for Pro for the previous 8 months. The lady I contacted in support simply told me to pay the outstanding balance. You would think if you missed a payment it would relate to simply that month that was missed not all subsequent months.
Utterly ridiculous that one missed payment can justify not providing the service (otherwise paid for in full) at all.
Basically if you find yourself in this situation you're actually better of deleting the account and resigning up again under a different email.
We really need to get our shit together in the EU on this sort of stuff, I was a paying customer purely out of sympathy but that sympathy dried up pretty quick with hostile customer service.
Re: Mistral 3 family of models released
#53Like how does 14B compare to Qwen30B-A3B?
(Which I think is a lot of people's goto or it's instruct/coding variant, from what I've seen in local model circles)
Re: Mistral 3 family of models released
#54Re: Mistral 3 family of models released
#55Earlier quoted context omitted.
Until there is a sustainable, profitable and moat-building business model for generative AI, the competition is not to have the best proprietary model, but rather to raise the most VC money to be well positioned when that business model does arise. Releasing a near stat-of-the-art open model instanly catapults companies to a valuation of several billion dollars, making it possible raise money to acquire GPUs and trai…
Explained well in this documentary [0]. [0] https://www.youtube.com/watch?v=BzAdXyPYKQo
Re: Mistral 3 family of models released
#56Earlier quoted context omitted.
Completely agree, that there are legitimate reasons to prefer comparison to e.g. deepeek models. But that doesn't change my point, we both agree that the comparisons would be extremely unfavorable.
> that the comparisons would be extremely unfavorable. Why should they compare apples to oranges? Ministral3 Large costs ~1/10th of Sonnet 4.5. They clearly target different users. If you want a coding assistant you probably wouldn't choose this model for various reasons. There is place for more than only the benchmark king.
Re: Mistral 3 family of models released
#57Europe's bright star has been quiet for a while, great to see them back and good to see them come back to Open Source light with Apache 2.0 licenses - they're too far from the SOTA pack that exclusive/proprietary models would work in their favor. Mistral had the best small models on consumer GPUs for a while, hopefully Ministral 14B lives up to their benchmarks.
All thanks to the US VCs that acutally have money to fund Mistral's entire business. Had they gone to the EU, Mistral would have gotten a miniscule grant from the EU to train their AI models.
Re: Mistral 3 family of models released
#58It's sad that they only compare to open weight models. I feel most users don't care much about OSS/not OSS. The value proposition is the quality of the generation for some use case. I guess it says a bit about the state of European AI
I also think most people do not consider open weights as OSS.
Re: Mistral 3 family of models released
#59Earlier quoted context omitted.
All thanks to the US VCs that acutally have money to fund Mistral's entire business. Had they gone to the EU, Mistral would have gotten a miniscule grant from the EU to train their AI models.
1. so what 2. asml
2. ASML was propped up by ASM and Philips, stepping in as "VCs"
Re: Mistral 3 family of models released
#60I use large language models in http://phrasing.app to format data I can retrieve in a consistent skimmable manner. I switched to mistral-3-medium-0525 a few months back after struggling to get gpt-5 to stop producing gibberish. It's been insanely fast, cheap, reliable, and follows formatting instructions to the letter. I was (and still am) super super impressed. Even if it does not hold up in benchmarks, it still out…
Does Mistral even have a Tool Use model? That would be awesome to have a new coder entrant beyond OpenAI, Anthropic, Grok, and Qwen.