Live data from Hacker News

Magistral — the first reasoning model by Mistral AI

mistral.ai

241–250 of 444 posts

Re: Magistral — the first reasoning model by Mistral AI

#241

Earlier quoted context omitted.

Are we sure more time butt in office equates to more productivity?

I think maybe we should completely switch to admitting this. Every extra second you sit in the (home)office adds to productivity, just not necessarily converting into market values, that can be inflated with hype. Also longer hours is not necessarily safe or sustainable. We only wish more time != more productivity because it's inconvenient in multiple ways if it were. We imagine a multiplier in there to balance the e…

> Every extra second you sit in the (home)office adds to productivity

I'm not sure I believe that. I think at some point the additional hours worked will ultimately decrease the output/unit of time and at some point that you'll reach a peak whereafter every hour worked extra will lead to an overall productivity loss.

Its also something that I think is extremely hard to consistently measure, especially for your typical office worker.

Re: Magistral — the first reasoning model by Mistral AI

#242
post #223
post #218

Earlier quoted context omitted.

too much thinking https://gist.github.com/gavi/b9985f730f5deefe49b6a28e5569d46...

My impression from running the first R1 release locally was that it also does too much thinking.

It does not do any thinking. It is a statistical model, just like the rest of them.

Re: Magistral — the first reasoning model by Mistral AI

#243

Earlier quoted context omitted.

probably some silly thing like "people should have more rights and protections"

Rights and protections that have benefited heavily from an economy built on the alliance with the US. If it weren't for American help and trade post-WW2, Europe would be a Belarusian backwater and is fast heading back in that direction. Countries like Greece, Italy, Spain, Portugal, etc. show the future of Europe as it slowly stagnates and becomes a museum that can't feed it's people. Even Germany that was once excel…

> Countries like Greece, Italy, Spain, Portugal

PIGS, really? Some of the top growing EU economies right now, which have turned their deficit around, show the future of a slowly stagnating Europe?

Re: Magistral — the first reasoning model by Mistral AI

#244

Earlier quoted context omitted.

A similar sentiment existed for a long time about Uber and now they're very profitable and own their market. It was worth the burn to capture the market. Who says OpenAI can't roll over to profitable at a stable scale? Conquer the market, hike the price to $29.95 (family account, no ads; $19.95 individual account with ads; etc etc). To say nothing of how they can branch out in terms of being the interaction point tha…

> It was worth the burn to capture the market. You cannot compare Uber to the AI market. They are too different. Uber captured the market because having three taxi services is annoying. But people are readily jumping between models using multi-model platforms. And nobody is significantly ahead of the pack. There is nothing that sets anyone apart aside from the rate at which they are burning capital. Any advantage is…

Three cab apps are a lot less annoying than three LLM apps each having their piece of your chats history.

The winner-take-all effect is a lot stronger with chat apps.

Re: Magistral — the first reasoning model by Mistral AI

#245
post #199

Earlier quoted context omitted.

If I work 1000 hours and you work 2000 hours in the same timeframe, but you outcompeted me and created 3x value, you are 1.5 times more productive. There's a numerator too.

How does the same exact person get more productive? You forgot the example I replied to? The only thing that changed were hours worked. In your example you change it to less hours worked with more output. You made it circular.

You can be more productive just because you're faster.

Magistral is amazingly impressive compared to ChatGPT 3.5. If it had come out two years ago we'd be saying Mistral is the clear leader. But it came out now.

Not saying they worked fewer hours, just that speed matters, and in some cases, up to a limit, working more hours gets your work done faster.

Re: Magistral — the first reasoning model by Mistral AI

#246
post #239

As a quick test of logical reasoning and basic Wikipedia-level knowledge, I asked Mistral AI the following question: A Brazilian citizen is flying from Sao Paulo to Paris, with a connection in Lisbon. Does he need to clear immigration in Lisbon or in Paris or in both cities or in neither city? Mistral AI said that "immigration control will only be cleared in Paris," which I think is wrong. After I pointed it to the W…

doing some reason.. uhh intuitioning i imagine brazil and portugal might have some sort of a visa-free deal going on in which case llama 4 might actually be right here?

Brazilians don't need a visa for Portugal, France, or any Schengen country. But everybody has to pass through immigration control (at least a passport check even if you don't need a visa) when entering the Schengen zone. My question was which country would that happen in.

Re: Magistral — the first reasoning model by Mistral AI

#247
post #173

Earlier quoted context omitted.

Because Europeans don't take smart risks. Because they over regulate. It's fascinating watching people circle back to this answer. Regulation and taxation reduces incentives. Lower incentives, means lower risk-taking. The fact this is still a lesson that needs to be debated is absurd.

Europeans also mostly don’t suffer from school shootings and generally don’t go bankrupt when they get cancer or just take an ambulance ride to a non-network hospital. Regulation is not all bad, besides the US has more of it than anybody else.

The mental gymnastics here are incredible. Do you really think the regulations inhibiting tech startup creation are the same ones that protect people when they get cancer or whatever?

Yes, the US has a lot of school shootings, but does anyone think loose gun regulations are why the US is strong on tech?

Re: Magistral — the first reasoning model by Mistral AI

#248

Earlier quoted context omitted.

With how amazing the first R1 model was and how little compute they needed to create it, I'm really wondering how the new R1 model isn't beating o3 and 2.5 Pro on every single benchmark. Magistral Small is only 24B and scores 70.7% on AIME2024 while the 32B distill of R1 scores 72.6%. And with majority voting @64 the Magistral Small manages 83.3%, which is better than the full R1. Since I can run a 24B model on a reg…

It's because DeepSeek was a fast copy. That was the easy part and it's why they didn't have to use so much compute to get near the top. Going well beyond o3 or 2.5 Pro is drastically more expensive than fast copy. China's cultural approach to building substantial things produces this sort of outcome regularly, you see the same approach in automobiles, planes, Internet services, industrial machinery, military, et al.…

I understand that the French are very innovative so why isn't their model SOTA ?

Re: Magistral — the first reasoning model by Mistral AI

#249

Earlier quoted context omitted.

With how amazing the first R1 model was and how little compute they needed to create it, I'm really wondering how the new R1 model isn't beating o3 and 2.5 Pro on every single benchmark. Magistral Small is only 24B and scores 70.7% on AIME2024 while the 32B distill of R1 scores 72.6%. And with majority voting @64 the Magistral Small manages 83.3%, which is better than the full R1. Since I can run a 24B model on a reg…

It's because DeepSeek was a fast copy. That was the easy part and it's why they didn't have to use so much compute to get near the top. Going well beyond o3 or 2.5 Pro is drastically more expensive than fast copy. China's cultural approach to building substantial things produces this sort of outcome regularly, you see the same approach in automobiles, planes, Internet services, industrial machinery, military, et al.…

[deleted]

Re: Magistral — the first reasoning model by Mistral AI

#250
post #168

Earlier quoted context omitted.

Can you please explain what your „real life situations“ are?

I use it as a personal assistant (so tool use integrated into calendar/todo/notes etc) often times using the multimodal aspect (taking a photo of a todo list, asking it to remind me to buy something from a picture). I also use it as a code completion tool in vscode, as well as a replacement for most basic google searches ("how does this syntax work", "what's the torch method for X") I use it for almost every interact…

Cool. What framework or program do you use to orchestrate this?
Post reply on HN