Live data from Hacker News

Magistral — the first reasoning model by Mistral AI

mistral.ai

371–380 of 444 posts

Re: Magistral — the first reasoning model by Mistral AI

#371

Earlier quoted context omitted.

I think usb-c and third party app stores are pretty cool

I think the government shouldn't be legislating that companies must use a specific USB connector. Realistically the legislation was only targeting Apple. If consumers want USB-C, then they can vote with their wallets and buy an Android, which is a reasonable alternative.

It used to be the case in Europe that you couldn't use a washing machine made for Sweden in Norway. Everything was different. Every country had its own standards too, which had to certify your products. It was openly for protectionistic reasons.

EU got rid of that. It only makes sense that they don't let private companies start all that crap up again. If states don't get to use artificial technological barriers as protectionism, certainly Apple shouldn't be allowed to either.

Re: Magistral — the first reasoning model by Mistral AI

#372
post #366

Earlier quoted context omitted.

Facebook and Google were never big in China, not even close.

Google had 31% market share in 2010 http://news.bbc.co.uk/2/hi/business/8455712.stm I haven't been able to find numbers for Facebook.

Don't have a source either but I was living in China back then and basically no one was using it. It was QQ and Renren.

Re: Magistral — the first reasoning model by Mistral AI

#373
post #30

Is the number of em-dashes in this marketing copy indicative of the kind of output that the model produces? If so, might want to tone it down a bit.

We don’t have em dashes as punctuation in French —- commas are usually used instead —- so we get overly excited about using them when we can —- everybody likes novelty.

Re: Magistral — the first reasoning model by Mistral AI

#374
post #302

Earlier quoted context omitted.

Interesting presumption about R1 25-01 being what's talked about, you knowledge cut-off does appear to know R1 update two weeks back was a thing, and that it even improved on function calling. Of course you have to pretend I meant the former, otherwise "they all have" doesn't entirely make sense. Not that it made total sense before either, but if I say your definition of "they" is laughably narrow, I suspect you will…

I don't know what you're talking about, partially because of poor grammar ("you knowledge cut-off does appear") and "presumption" (this was front and center on their API page at r1 release, and its in the r1 update notes). I sort of stopped reading after there because I realized you might be referring to me having a "knowledge cut-off", which is bizarre and also hard to understand, and it's unlikely to be particularl…

> you might be referring to me having a "knowledge cut-off"

Don't forget I also referred to you having "hallucination". In retrospect, likening your logical consistency to an LLM was premature, because not even gpt-3.5 era models could pull off a gem like:

> You: to your Q about why no one can compete with DeepSeek R1 25-01 blah blah blah

>> Me: ...why would you presume I was talking about 25-01 when 28-05 exists and you even seem to know it?

>>> You: this was front and center on their API page!

Riveting stuff. Few more digs about poor grammar and how many times you stopped reading, and you might even sell the misdirection.

Re: Magistral — the first reasoning model by Mistral AI

#375

Earlier quoted context omitted.

Because we do not have a complete understanding of human neurons. How are we supposed to accurately model something we cannot directly observe?

Just because you don't know how does not mean that we can't.

Prove it, then.

Re: Magistral — the first reasoning model by Mistral AI

#376

Earlier quoted context omitted.

I think maybe we should completely switch to admitting this. Every extra second you sit in the (home)office adds to productivity, just not necessarily converting into market values, that can be inflated with hype. Also longer hours is not necessarily safe or sustainable. We only wish more time != more productivity because it's inconvenient in multiple ways if it were. We imagine a multiplier in there to balance the e…

> Every extra second you sit in the (home)office adds to productivity I'm not sure I believe that. I think at some point the additional hours worked will ultimately decrease the output/unit of time and at some point that you'll reach a peak whereafter every hour worked extra will lead to an overall productivity loss. Its also something that I think is extremely hard to consistently measure, especially for your typica…

Here you go https://cs.stanford.edu/people/eroberts/cs181/projects/crunc...

Re: Magistral — the first reasoning model by Mistral AI

#377

Earlier quoted context omitted.

Nobody should be ever using ollama, for any reason. It literally only makes everything worse and more convoluted with zero benefits.

Could you elaborate?

ollama is just a wrapper for llama.cpp that adds insane defaults.

Just use llama.cpp directly.

Re: Magistral — the first reasoning model by Mistral AI

#378
post #30

Is the number of em-dashes in this marketing copy indicative of the kind of output that the model produces? If so, might want to tone it down a bit.

This meme that humans don’t use em dashes needs to die. It’s an extremely useful tool in writing and I’ve been using it for decades.

I love a good em-dash, but this page overuses them (nearly 1:1 ratio of em-dashes to commas!) and puts them in places where they just do not belong.

Re: Magistral — the first reasoning model by Mistral AI

#379

Earlier quoted context omitted.

Indeed, and with the technology plateau-ing, being 6-12 months late with less debt is just long term thinking. Also, Europe being in the race is a big deal for consumers.

>with the technology plateau-ing People were claiming that since year 2022. Where's the plateau?

There's frequent discussions about how sonnet-3.5 is in the same ballpark or even outperforms sonnet-3.7 and 4.0, for example.

Re: Magistral — the first reasoning model by Mistral AI

#380

Earlier quoted context omitted.

It's not better than full R1; Mistral is using misleading benchmarks. The latest version of R1, R1-0528, is much better: 91.4% on AIME2024 pass@1. Mistral uses the original R1 release from January in their comparisons, presumably because it makes their numbers look more competitive. That being said, it's still very impressive for a 24B. I'm really wondering how the new R1 model isn't beating o3 and 2.5 Pro on every s…

It may not have been intentionally misleading. Some benchmarks can take a lot of horsepower and time to run. Their preparation for release likely was done well in advance of the model release before the new deepseek r1 model had even been available to test.

AIME24, etc are pretty cheap to run using any DeepSeek API. Regardless, they didn't even run the benchmarks for R1 themselves, they just republished DeepSeek's published numbers from January. They could have published the ones from May, but chose not to.
Post reply on HN