Live data from Hacker News

Leanstral 1.5

docs.mistral.ai

31–40 of 158 posts

Re: Leanstral 1.5

#31
post #8
post #4

Earlier quoted context omitted.

I like the models for creative writing. They have a distinct voice that is different from the other llms.

I made a game ( https://prose-or-con.com ) where you pick whether writing is AI or human. Mistral is a bonkers weird writer. So weird I fell for it a couple of times because I thought, "No way a model writes this weird." Not, like, incorrect grammar or spelling or anything, just...off-kilter. Kinda sassy.

needs a leaderboard of models most often mistaken for humans.

Re: Leanstral 1.5

#32
post #8

Earlier quoted context omitted.

I made a game ( https://prose-or-con.com ) where you pick whether writing is AI or human. Mistral is a bonkers weird writer. So weird I fell for it a couple of times because I thought, "No way a model writes this weird." Not, like, incorrect grammar or spelling or anything, just...off-kilter. Kinda sassy.

needs a leaderboard of models most often mistaken for humans.

Yes, it's on the todo list, but I need more data. Only a half dozen people have played it and submitted a score. I'm storing the hashes of passages people got right and wrong so I can make exactly that chart at some point. I think both "the most human-like AI" and "the most AI-like human" are both interesting pieces of data, but I don't know either yet.

Re: Leanstral 1.5

#33
Got curious, sign up, add money to account, try to use. Can't, it's a labs model. Fine, let's enable labs. Can't, unspecified error. Fine, lets contact customer support as instructed, can't no customer support, just a half-assed FAQ, that seems vibe-coded and searched poorly, totally irrelevant answers coming up for all queries tried. Then it hit me:

If AI makes good customer support, then why does no AI company use theirs to provide customer support?

Re: Leanstral 1.5

#34
post #17

Real talk, does anyone use anything from Mistral because it performs the best, by whatever secular metric of your choosing? Or is it only used "because EU"? Just focus on answering the question. I wonder if anyone has observed it perform better on any objective metric in any rigorous setting.

For writing and languange learning it's very decent, especially Mistral Large. The pricing is very good too. I really like the consistently low time to first token and good token per second. Claude, especially in the past, would be very inconsistent, often with outages. Mistral mostly just always works and is very fast. Technical questions are unfortunately hit or miss. I'm lately pretty much always using a system pr…

vibe has improved _a lot_ during the past few months, fyi.

The new Mistral Medium 3.5 is also a big improvement over devstral-2

Re: Leanstral 1.5

#35

Got curious, sign up, add money to account, try to use. Can't, it's a labs model. Fine, let's enable labs. Can't, unspecified error. Fine, lets contact customer support as instructed, can't no customer support, just a half-assed FAQ, that seems vibe-coded and searched poorly, totally irrelevant answers coming up for all queries tried. Then it hit me: If AI makes good customer support, then why does no AI company use…

Because that AI will either expose their business or it will be so nerfed it’s useless.

Re: Leanstral 1.5

#36

Real talk, does anyone use anything from Mistral because it performs the best, by whatever secular metric of your choosing? Or is it only used "because EU"? Just focus on answering the question. I wonder if anyone has observed it perform better on any objective metric in any rigorous setting.

>Just focus on answering the question.

Are you trying to instruct me like an LLM?

Re: Leanstral 1.5

#38

Real talk, does anyone use anything from Mistral because it performs the best, by whatever secular metric of your choosing? Or is it only used "because EU"? Just focus on answering the question. I wonder if anyone has observed it perform better on any objective metric in any rigorous setting.

I use it as my workhorse for coding and general chat questions, because it's good enough 80% of the time, and indeed it's french/european (with heavy US capital tho...).

We complain too much about not having enough major competitors in the IT space, to not support a burgeoning one even if it's less powerful than SOTA labs

Re: Leanstral 1.5

#39

Earlier quoted context omitted.

needs a leaderboard of models most often mistaken for humans.

Yes, it's on the todo list, but I need more data. Only a half dozen people have played it and submitted a score. I'm storing the hashes of passages people got right and wrong so I can make exactly that chart at some point. I think both "the most human-like AI" and "the most AI-like human" are both interesting pieces of data, but I don't know either yet.

try posting it on r/localllama and r/sillytavernai

Re: Leanstral 1.5

#40

Got curious, sign up, add money to account, try to use. Can't, it's a labs model. Fine, let's enable labs. Can't, unspecified error. Fine, lets contact customer support as instructed, can't no customer support, just a half-assed FAQ, that seems vibe-coded and searched poorly, totally irrelevant answers coming up for all queries tried. Then it hit me: If AI makes good customer support, then why does no AI company use…

This isn’t the first time. I’m amazed at how they manage to fumble releases over and over …
Post reply on HN