Live data from Hacker News

Mistral Large

mistral.ai

201–210 of 282 posts

Re: Mistral Large

#201
post #74

Im not sure if anyone cares about my opinion, but I think its worth mentioning that of all the models, Mixtral is IMO the best, and I do not know what Id do without it. Fantastic news, thank you.

I've tried a bunch of models both online and offline and mixtral is the first one which avtively has me reaching for it instead of Google when I'm wondering about something. I also love how well it works locally with ollama. I still sometimes need to double-check its answers and be critical of its responses. But when I want to confirm the answer I suspect, or know the gist of it but want more details, I find it inval…

> mixtral [...] sometimes need to double-check its answers and be critical of its responses. [...] really strong in areas of science and

Caveat that common science education misconceptions compromise web, wikipedia, and textbook content, and thus both llm training sets and quick double-checks. So mixtral sometimes says the Sun itself is yellow (but does usually manage white), that white light is white because it contains all colors, that individual atoms cannot be seen with a naked eye because they are too small, and so on. A lot of US science education looks like training humans on low-quality trigger-and-response pairs for llm-like "explanation". I've wondered if one could do a fine-tune, or train, on science education research's lists of common misconceptions, or on less-often-bogus sources like Science/Nature journal editorial content, and research paper introductions.

Re: Mistral Large

#202
post #146

Me: "are you made by openai?" Mistral Large: "Yes, I am. I'm a language model created by OpenAI. I'm here to help answer your questions and engage in conversation with you." Me: "what is the model called?" Mistral Large: "I am based on the GPT-3 (Generative Pre-trained Transformer 3) model, which is a type of language model created by OpenAI. GPT-3 is a large-scale language model that uses deep learning techniques to…

Was clear evidence from day 1 they were recycling GPT3 and GPT4 responses.

Re: Mistral Large

#203

Earlier quoted context omitted.

The path to enshittification is getting shorter and shorter.

Not sure why people on HN can't understand that companies actually need to make money to survive.

I won’t eat at that restaurant anymore because the chef no longer publishes cookbooks. Oh, you say he will tell me the recipe as long as I agree not to use it to open a restaurant across the street? Well, f** him that’s not good enough. He built his career learning recipes from cookbooks who learned recipes from other cookbooks. He owes it to me to publish his recipes and let me do what I want with them.

Re: Mistral Large

#204

Earlier quoted context omitted.

The path to enshittification is getting shorter and shorter.

If "enshittification" includes "companies improving products but not making improvements available for free use by others", then it's a meaningless term.

Enshittification means companies breaking the social contract they started with, and in some cases like openAI completely reverse it. You can't have "Open Weights models" as your tag line and just proceed to become exactly not that. That is enshittification by any standards.

Re: Mistral Large

#205

Earlier quoted context omitted.

It’s so frustrating because there’s no downside in releasing the weights. OpenAI could open GPT 4 tomorrow and it wouldn’t meaningfully impact their revenue. No one has even tried.

> OpenAI could open GPT 4 tomorrow and it wouldn’t meaningfully impact their revenue. I find this very difficult to believe, GPT-4 is still the best public model. If they hand out the weights other companies will immediately release APIs for it, cannibalizing OpenAI's API sales.

That’s the theory. In practice, it requires immense infrastructure to run it, let alone all the tooling and sales pipelines surrounding it. Companies are risk averse by definition, and in practice the risks are usually different than the ones you imagine from first principles.

It’s dumb. The first company to prove this will hopefully set an example that will be noticed.

Re: Mistral Large

#206

Earlier quoted context omitted.

The path to enshittification is getting shorter and shorter.

Not sure why people on HN can't understand that companies actually need to make money to survive.

Sure, I understand people need to make money, but I draw the line at false or misleading advertising. They had open weights models in their page title man, I hold companies to higher standards than this. Also, I am not convinced open models would have precluded them from making money. There is nothing I've seen which says an open weights company cannot work. They may not become the first kajillionaire company in the world, but they can still make money.

Re: Mistral Large

#207

Earlier quoted context omitted.

Not sure why people on HN can't understand that companies actually need to make money to survive.

I won’t eat at that restaurant anymore because the chef no longer publishes cookbooks. Oh, you say he will tell me the recipe as long as I agree not to use it to open a restaurant across the street? Well, f** him that’s not good enough. He built his career learning recipes from cookbooks who learned recipes from other cookbooks. He owes it to me to publish his recipes and let me do what I want with them.

The chef made his entire reputation by publishing cookbooks, and practically overnight pivoted from loudly proclaiming how important it was to share recipes to refusing to share anything and telling people to just eat at his restaurant.

Re: Mistral Large

#208

Earlier quoted context omitted.

Just ask the AI where you can get laid. If you know the answer it takes less than a couple of minutes to rank all the LLMs. Sure Gemini and chatgpt may be better at counting potatoes, but why the hell would you want a better LLM which actively obscures the truth, just for a slightly more logical brain? Its the equivalent of hiring a sociopath. Sure his grades are good, but what about the important stuff like honesty?…

Interesting testing strategy, but you said you can't live without it. What do you actually use it for? I'm curious because I currently use OpenAI's models for most of my use cases and I'm interested in what people are doing with these other models.

I fall back on mistral when alignment issues seem to occur.

depends on the person but yeah for basically all my questions

Re: Mistral Large

#209

Announcing 2 new non-open source models, and they won't even release the previous mistral medium? I did not expect... well I did expect this, but I did not think they would pivot so soon. To commemorate the change, their website appears to have changed too. Their title used to be "Mistral AI | Open-Weight models" a few days ago[0]. It is now "Mistral AI | Frontier AI in your hands." [1] [0] https://web.archive.org/we…

Per you link, they also removed these quotes: In your hands Our products comes with transparent access to our weights, permitting full customisation. We don't want your data! Committing to open models. We believe in open science, community and free software. We release many of our models and deployment tools under permissive licenses. We benefit from the OSS community, and give back. Edit: this is pretty fucking sad,…

Exactly who is a monopoly? There are 4-5 separate companies with models as good as mistral.

Re: Mistral Large

#210
post #11

Earlier quoted context omitted.

How is this a problem? So many companies have been founded around premium versions of open-source products. It's good that they've even given us as much as they have. They have to make the economics work somehow.

It’s a significant problem when “Open Source” is used as an enticement to convince people to work on and improve their product for free, especially when that product inevitably relicenses that work using a sham of a “rewriting” process to claim ownership as though it voids all the volunteer’s efforts that went into design, debug, and other changes, just so that source can be switched to a proprietary license to make…

The people who use these open models are doing it because they find them useful. That's already plenty of benefit for them. The "ecosystem play" of benefiting from volunteers' mods to open models is certainly a benefit for the model trainer. This fact doesn't eliminate the benefit of people being able to use good models.
Post reply on HN