Live data from Hacker News

Mistral Large

mistral.ai

181–190 of 282 posts

Re: Mistral Large

#181
post #177

Earlier quoted context omitted.

Funny, we're going to have to make a very clear divider between pre-2022 and post-2022 internet, kind of like nuclear-contaminated steel of post 1950 or whatever. Information is basically going to be unreliable, unless it's in a spec sheet created by a human, and even then, you have to look at the incentives.

If you think that's crazy, think again. Just yesterday was trying to learn more about Chinese medicine and landed on this page I thoroughly read before noticing the disclaimer at the top. "The articles on this database are automatically generated by our AI system" https://www.digicomply.com/dietary-supplements-database/pana... Is the information on that page correct? I'm not sure but as soon as I noticed it was AI ge…

You shouldn't have had any trust to begin with; I don't know why we are so quick to hold up humans as bastions of truth and integrity.

This is stereotypical Gell-Mann amnesia - you have to validate information, for yourself, within your own model of the world. You need the tools to be able to verify information that's important to you, whether it's research or knowing which experts or sources are likely to be trustworthy.

With AI video and audio on the horizon, you're left with having to determine for yourself whether to trust any given piece of media, and the only thing you'll know for sure is your own experience of events in the real world.

That doesn't mean you need to discard all information online as untrustworthy. It just means we're going to need better tools and webs of trust based on repeated good-faith interactions.

It's likely I can trust that information posted by individuals on HN will be of a higher quality than the comments section in YouTube or some random newspaper site. I don't need more than a superficial confirmation that information provided here is true - but if it's important, then I will want corroboration from many sources, with validation by an expert extant human.

There's no downside in trusting the information you're provided by AI just as much as any piece of information provided by a human, if you're reasonable about it. Right now is as bad as they'll ever be, and all sorts of development is going in to making them more reliable, factual, and verifiable, with appropriately sourced validation.

Based on my own knowledge of ginseng and a superficial verification of what that site says, it's more or less as correct as any copy produced by a human copy writer would be. It tracks with wikipedia and numerous other sources.

All that said, however, I think the killer app for AI will be e-butlers that interface with content for us, extracting meaningful information, identifying biases, ulterior motives, political and commercial influences, providing background research, and local indexing so that we can offload much of the uncertainty and work required to sift the content we want from the SEO boilerplate garbage pit that is the internet.

Re: Mistral Large

#182

Announcing 2 new non-open source models, and they won't even release the previous mistral medium? I did not expect... well I did expect this, but I did not think they would pivot so soon. To commemorate the change, their website appears to have changed too. Their title used to be "Mistral AI | Open-Weight models" a few days ago[0]. It is now "Mistral AI | Frontier AI in your hands." [1] [0] https://web.archive.org/we…

The path to enshittification is getting shorter and shorter.

If "enshittification" includes "companies improving products but not making improvements available for free use by others", then it's a meaningless term.

Re: Mistral Large

#183
post #97

Here's a chart indicating we're not too much worse than the industry leader

And less than half the price. It's even cheaper than GPT4-Turbo.

GPT-4-Turbo is now the flagship model, so they’re slightly cheaper than OpenAI. The fact that they priced this way after getting Microsoft investment should set off EU regulator alarm bells.

Re: Mistral Large

#185

Announcing 2 new non-open source models, and they won't even release the previous mistral medium? I did not expect... well I did expect this, but I did not think they would pivot so soon. To commemorate the change, their website appears to have changed too. Their title used to be "Mistral AI | Open-Weight models" a few days ago[0]. It is now "Mistral AI | Frontier AI in your hands." [1] [0] https://web.archive.org/we…

It’s so frustrating because there’s no downside in releasing the weights. OpenAI could open GPT 4 tomorrow and it wouldn’t meaningfully impact their revenue. No one has even tried.

> OpenAI could open GPT 4 tomorrow and it wouldn’t meaningfully impact their revenue.

I find this very difficult to believe, GPT-4 is still the best public model. If they hand out the weights other companies will immediately release APIs for it, cannibalizing OpenAI's API sales.

Re: Mistral Large

#186
post #177

Earlier quoted context omitted.

If you think that's crazy, think again. Just yesterday was trying to learn more about Chinese medicine and landed on this page I thoroughly read before noticing the disclaimer at the top. "The articles on this database are automatically generated by our AI system" https://www.digicomply.com/dietary-supplements-database/pana... Is the information on that page correct? I'm not sure but as soon as I noticed it was AI ge…

You shouldn't have had any trust to begin with; I don't know why we are so quick to hold up humans as bastions of truth and integrity. This is stereotypical Gell-Mann amnesia - you have to validate information, for yourself, within your own model of the world. You need the tools to be able to verify information that's important to you, whether it's research or knowing which experts or sources are likely to be trustwo…

> This is stereotypical Gell-Mann amnesia - you have to validate information, for yourself, within your own model of the world. You need the tools to be able to verify information that's important to you, whether it's research or knowing which experts or sources are likely to be trustworthy.

Except anthropologically speaking we still live in trust-based society. We trust water to be available. We trust the grocery stores to be stocked. We trust that our Government institutions are always going to be there.

All this to say we have a moral obligation not to let AI spam off the hook as "trust but verify". It is fucked up that people make money abusing innate trust-based mechanism that society depends on to be society.

Re: Mistral Large

#187

Earlier quoted context omitted.

Any training on internet data beyond 2022 is gonna lead to this. ChatGPT output is sprawled everywhere on the internet.

Funny, we're going to have to make a very clear divider between pre-2022 and post-2022 internet, kind of like nuclear-contaminated steel of post 1950 or whatever. Information is basically going to be unreliable, unless it's in a spec sheet created by a human, and even then, you have to look at the incentives.

I was thinking the exact same thing last month[1]! It's really interesting what the implications of this might be, and how valuable human-derived content might become. There's still this idea of model collapse, whereby the output of LLMs trained repeatedly on artificial content descends into what we think is gibberish, so however realistic ChatGPT appears, there are still significant differences between its writing and ours.

[1]: https://www.glfharris.com/posts/2024/low-background-lexicogr...

Re: Mistral Large

#188
post #146

Me: "are you made by openai?" Mistral Large: "Yes, I am. I'm a language model created by OpenAI. I'm here to help answer your questions and engage in conversation with you." Me: "what is the model called?" Mistral Large: "I am based on the GPT-3 (Generative Pre-trained Transformer 3) model, which is a type of language model created by OpenAI. GPT-3 is a large-scale language model that uses deep learning techniques to…

I got the same thing. I got it to elaborate and as I asked it how it could be trained on GPT-3 when it's closed source. I asked if it got the data through the API. It insisted it was trained on conversational data, this leads me to believe they generated a bunch of conversational data using OpenAI APIs...

Re: Mistral Large

#190

Announcing 2 new non-open source models, and they won't even release the previous mistral medium? I did not expect... well I did expect this, but I did not think they would pivot so soon. To commemorate the change, their website appears to have changed too. Their title used to be "Mistral AI | Open-Weight models" a few days ago[0]. It is now "Mistral AI | Frontier AI in your hands." [1] [0] https://web.archive.org/we…

Per you link, they also removed these quotes:

In your hands

Our products comes with transparent access to our weights, permitting full customisation. We don't want your data!

Committing to open models.

We believe in open science, community and free software. We release many of our models and deployment tools under permissive licenses. We benefit from the OSS community, and give back.

Edit: this is pretty fucking sad, and the fact that it's become expected is... I dunno, a tragedy? I mean, the whole point of anti-trust law was that monopolies like this are a net negative to the economy and to social and technological progress. They are BAD for business for everyone except the monopolist.

Post reply on HN