Live data from Hacker News

The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

thesequence.substack.com

101–110 of 527 posts

Re: The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

#101
post #97
post #23

Earlier quoted context omitted.

They clearly expected the leak, they distributed it very widely to researchers. The important thing is the licence, not the access: you are not allowed to use it for commercial purpose.

How could Meta ever find out your private business is using their model without a whistleblower? It's practically impossible.

I think you can make that argument for all behind-the-scenes commercial copyright infringement, surely?

Re: The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

#102
post #48
post #45

Earlier quoted context omitted.

I am shocked that it speaks the way it does when it was trained on random stuff it doesn’t have rights to. They say they trained it on databases they had bought access to etc. And it seems that way. Because how does ChatGPT: 1. Do what you ask instead of continuing your instructions? 2. Use such nice and helpful language as opposed to just random average of what people say? 3. And most of all — how does it have a str…

It is steered by RLHF to give helpful, nice, structured continuations. it was totally trained on random text they never paid a dime for.

Can you please elaborate and answer 1,2,3 directly — I’d love to find out more. Maybw links to the techniques

Re: The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

#104

Earlier quoted context omitted.

I would assume so. Meta’s ML/AI team is very strong, but they probably don’t have a comparable product offering to ChatGPT ready for public use. So instead, they bought themselves some time by letting the open source community run wild with a lesser model and eat into OpenAI’s moat.

They didn't leak it. Someone else did.

I dont think theyre saying Meta AI leaked it, but they anticipated someone else will and still went ahead with it as they wanted the consequences.

Re: The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

#105
post #24

Earlier quoted context omitted.

There's been quite a few different iterations of ChatGPT and bing with different behaviours in this regard: it depends somewhat on the base GPT version, the fine-tuning, and the prompt. Bing very famously at one point was extremely passive aggressive when challenged on basically anything. And while there's nothing intrinsic to the structure and training goals of LLMs which directs them towards more structured reasoni…

> Bing very famously at one point was extremely passive aggressive when challenged on basically anything. It still wasn't an indication of how confident it "felt" with its answers. It was just role-playing a more confident and aggressive chat bot than ChatGPT does.

Outside of the roleplay, there's no "it". The thing that actually interacts with the user is the persona. But for the same reason, it doesn't matter what the underlying LLM "really thinks" - it'll be the persona that'll try to con you, or write broken or malicious code.

Re: The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

#106
post #97
post #23

Earlier quoted context omitted.

They clearly expected the leak, they distributed it very widely to researchers. The important thing is the licence, not the access: you are not allowed to use it for commercial purpose.

How could Meta ever find out your private business is using their model without a whistleblower? It's practically impossible.

You can just ask if there is no output filtering

Re: The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

#107

Earlier quoted context omitted.

I would assume so. Meta’s ML/AI team is very strong, but they probably don’t have a comparable product offering to ChatGPT ready for public use. So instead, they bought themselves some time by letting the open source community run wild with a lesser model and eat into OpenAI’s moat.

They didn't leak it. Someone else did.

It's extremely common for a "leak" to actually be fully intentional, but the organization in question just wants plausible deniability to mitigate legal/political/popular blowback.

In order to preserve plausible deniability, the leak will look genuine in all aspects that are easy to simulate. "Someone else did it" is easy to simulate. A better gauge would be to see if anyone is caught and punished. If so, it was probably a real leak.

Re: The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

#108
post #88

Earlier quoted context omitted.

I thought the copyright office said that ai generated material isn’t copyrighted?

You’re correct. US law states that intellectual property can be copyrighted only if it was the product of human creativity, and the USCO only acknowledges work authored by humans at present. Machines and generative AI algorithms, therefore, cannot be authors, and their outputs are not copyrightable.

Sorry for asking a stupid question.

How much Theseus do I need to ship before I can copyright it as my own? Is there some threshold for how much of an AI generated work needs to be modified by "human creativity" prior to it being copyrightable?

Re: The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

#109
post #43

> OpenAI published a detailed blog post outlining some of the principles used to ensure safety in their models. The post emphasize in areas such as privacy, factual accuracy Am I the only one amused by the phrase “factual accuracy”? How many stories have we read like the one where it tries to ghost light the guy that this year is actually last year. “Oh, your phone must be wrong too, because there is no way I could b…

I hold a more charitable interpretation. We (the public) have found an important bug in the system, ie. GPT can lie (or "hallucinate"), even if you try to convince it not to lie. The bug is definitely lowering the usefulness of their product, as well as the public option about it. But I'll let the programmer who has never coded a bug cast the first stone. I wouldn't be surprised if they're scrambling internally to mi…

Meet my darling https://gist.github.com/firatsarlar/5f25ebcc4c33ec484e6cd95b... We need to - pure - clear non commercial, non owned AI first. It -hope lets say-, no, most of us still guessing, sniffin, ... suspicions abut who own GPT -LLM we need to abstract-, Who owns this - so called - open source product? MS - with the power of NVIDA's - ? OpenAI people? Why cant we talk our intentions, or whose product, or not to generalize it As LLMs. Because, lets say public, - and our ignore our delusions about the thing we made-, because, we're in a delusion, what is this ? People could try to eat and sue some, do we - purely - care for public. I need to skip my own experience, but, yes, if you ready to be delusional , you could easily make LLM your girlfriend, a liar, ... No, we - software community - need to stop and see what we have. We name it, public shape it, some abuse it for money that it made me ... , and sue some ... What is the end product ? A constant, machine - lets say system -, working to make this happen. Justice, we - devs-, world, ... money, people, puclic , we 're spendng time and resources a thing. Some claim its mine, yes, we need a concencuss ? No. Open source is the only thing if we - pure - lets say. I do not want to waste my time to figure out poke FB's LLM, or -pleae Open - AI girl. Lets learn to share, at least for us -devs- I'm really tired to think commercial facts of this thing. I'm really tired to this game we play. MS, say I have it, and make it yours. Boy - sorry open AI- I do not know you, is it yours, MS's , FB, they somehow managed to make a thing - inter-national, in Paris - Skip gaslighting, delusions what you own ?

Are we really there ?

Re: The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

#110
post #4

This makes is sound as if the Stanford and Berkeley teams also benefited from the leak, whereas I doubt they didn't have official access. So Alpaca/Vicuna/Koala projects would have probably happened anyway. The leak helped with popularity and demand and also somewhat positive PR for Meta, which makes me think they do not mind the leak that much.

Right. Most of the fine-tuned models we've seen so far have been by university teams.

Meta is not being very selective here. I applied for the download myself and got the links after two days (using a university email address).

Post reply on HN