Live data from Hacker News

Magistral — the first reasoning model by Mistral AI

mistral.ai

261–270 of 444 posts

Re: Magistral — the first reasoning model by Mistral AI

#261

As a quick test of logical reasoning and basic Wikipedia-level knowledge, I asked Mistral AI the following question: A Brazilian citizen is flying from Sao Paulo to Paris, with a connection in Lisbon. Does he need to clear immigration in Lisbon or in Paris or in both cities or in neither city? Mistral AI said that "immigration control will only be cleared in Paris," which I think is wrong. After I pointed it to the W…

I think Gemini's answer (2.5 Flash) is impressive

----

Since both Portugal and France are part of the Schengen Area, and a Brazilian citizen generally does not need a visa for short stays (up to 90 days in any 180-day period) in the Schengen Area, here's how immigration will work:

Lisbon: The Brazilian citizen will need to clear immigration in Lisbon. This is because Lisbon is the first point of entry into the Schengen Area. At this point, their passport will be stamped, and they will be officially admitted into the Schengen Zone.

Paris: Once they have cleared immigration in Lisbon, their flight from Lisbon to Paris is considered a domestic flight within the Schengen Area. Therefore, they will not need to clear immigration again in Paris.

Important Note: While Brazilians currently enjoy visa-free travel, the European Travel Information and Authorization System (ETIAS) is expected to become mandatory by late 2026. Once implemented, Brazilian citizens will need to obtain this electronic authorization before their trip to Europe, even for visa-free stays. However, this is a pre-travel authorization, not a visa in the traditional sense, and the immigration clearance process at the first point of entry would remain the same.

Re: Magistral — the first reasoning model by Mistral AI

#262
post #78

Earlier quoted context omitted.

Sorry this has nothing to do with the point you're making but I've literally never seen anyone use the word 'fluencers in place of influencers lol.

Me neither and it's not much shorter. I think fluzies could work better.

I propose effluencers.

Re: Magistral — the first reasoning model by Mistral AI

#263

I wished the charts included Qwen3, the current SOTA in reasoning. Qwen3-4B almost beats Magistral-22B on the 4 available benchmarks, and Qwen3-30B-A3B is miles ahead.

Is there a popular benchmark site people use? Becaues I had to test all these by hand and `Qwen3-30B-A3B` still seems like the best model I can run in that relative parameter space (/memory requirements).

- https://livebench.ai/#/ + AIME + LiveCodeBench for reasoning

- MMLU-Pro for knowledge

- https://lmarena.ai/leaderboard for user preference

We only got Magistral's GPQA, AIME & livecodebench so far.

Re: Magistral — the first reasoning model by Mistral AI

#264
post #166

Earlier quoted context omitted.

Hi Simon, What's the huge difference between the two pelicans riding bicycles? Was one running locally the small version vs the pretty good one running the bigger one thru the API? Thanks, Morgan

Ollama doesn't like proper naming for some reason, so `ollama pull magistral:latest` lands you with the q4_K_M version (currently, subject to change). Mistral's API defaults to `magistral-medium-2506` right now, which is running with full precision, no quantization.

Not only the quantization, but what’s available via ollama is magistral-small (for local inference), not the -medium variant.

Re: Magistral — the first reasoning model by Mistral AI

#265
post #109

Earlier quoted context omitted.

24B is the size of the Small opensourced model. The Medium model is bigger (they don't seem to disclose its size) and still gets beaten by Deepseek R1

Mistral Large is 123b so one can probably assume that medium is between 24b and 123b, also Mistral 3.1 is by a wide margin my go-to model in real life situations. Benchmarks absolutely don't tell the whole story, and different models have different use cases.

It's a 70b model, Medium 2 was 70b.

https://xcancel.com/arthurmensch/status/1920136871461433620#...

Re: Magistral — the first reasoning model by Mistral AI

#266
post #213

Earlier quoted context omitted.

Nobody is claiming the US has less mass shootings. It's just pointless whataboutism in a conversation (economic strategy) that has nothing to do with it.

Ah good, I thought you were trying to imply there is an equivalent problem in the EU. Which would seem to be intentionally dense of course.

[deleted]

Re: Magistral — the first reasoning model by Mistral AI

#267
post #223

Earlier quoted context omitted.

My impression from running the first R1 release locally was that it also does too much thinking.

It does not do any thinking. It is a statistical model, just like the rest of them.

What are we doing when we think?

Re: Magistral — the first reasoning model by Mistral AI

#268
Etymological fun: both "mistral" and "magistral" mean "masterly."

Mistral comes from Occitan for masterly, although today as far as I know it's only used in English when talking about mediterranean winds.

Magistral is just the adjective form of "magister," so "like a master."

If you want to make a few bucks, maybe look up some more obscure synonyms for masterly and pick up the domain names.

Re: Magistral — the first reasoning model by Mistral AI

#269
post #223

Earlier quoted context omitted.

My impression from running the first R1 release locally was that it also does too much thinking.

It does not do any thinking. It is a statistical model, just like the rest of them.

"Thinking" is a term of art referring to the hidden/internal output of "reasoning" models where they output "chain of thought" before giving an answer[1]. This technique and name stem from the early observation that LLMs do better when explicitly told to "think step by step"[2]. Hope that helps clarify things for you for future constructive discussion.

[1] https://arxiv.org/html/2410.10630v1

[2] https://arxiv.org/pdf/2205.11916

Re: Magistral — the first reasoning model by Mistral AI

#270

Earlier quoted context omitted.

Are we sure more time butt in office equates to more productivity?

$89,000 GDP per capita vs $46,000 rather proves the point about productivity per butt. US office workers are extraordinarily productive in terms of what their work generates (thanks to numerous well understood things like the outsized US scaling abilities). Measuring beyond that is very difficult due to the variance of every business.

> $89,000 GDP per capita vs $46,000 rather proves the point about productivity per butt.

So if I work 24h/day in a farm in Afghanistan, I should earn more than software developers in the Silicon Valley (because I'm pretty sure that they sleep)? Is that how you say GDP works?

Post reply on HN