Live data from Hacker News

The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

thesequence.substack.com

121–130 of 527 posts

Re: The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

#121

Earlier quoted context omitted.

They didn't leak it. Someone else did.

They have tacitly endorsed the leak. https://github.com/facebookresearch/llama/pull/73#issuecomme...

Only because publicly visible actions are worse for them

People have gotten DMCA takedown requests from them over Llama repositories

Re: The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

#122

I've spent an embarassing amount of time since the llamas leaked playing with them, the tools to run them, and writing wrappers for them. They are technically alternatives in the sense that they're incomparably better chat bots than anything in the past. But at least for the 30B and under versions (65B is too big for me to run), no matter what fine tuning is done (alpaca, gpt4all, vicuna, etc), the llamas themselves…

Your conclusion seems not to be warranted since you haven't tried out the 65B model.

Re: The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

#123
post #56

Earlier quoted context omitted.

>Everything it reads is mapped into language, not concept space Umm I'm pretty sure it's discovered concepts through compressing text - it seems perfectly capable of generalizing concepts

> it seems perfectly capable of generalizing concepts How would you support that perception?

With hope and living? It is a dream come true for people. An abstract perception of a knowledge, is like sniffing a rose. It feels, yes, I get there. This 40.000 pages book, woow, I'll make time to live it or sniff another daisy?!

Re: The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

#124

I've spent an embarassing amount of time since the llamas leaked playing with them, the tools to run them, and writing wrappers for them. They are technically alternatives in the sense that they're incomparably better chat bots than anything in the past. But at least for the 30B and under versions (65B is too big for me to run), no matter what fine tuning is done (alpaca, gpt4all, vicuna, etc), the llamas themselves…

I’ve got access to 4 and it’s a huge leap up from 3.5 - much more subtlety in the response, less hallucinations, less hitting a brick wall, but all of it adding up to a giant leap.

Re: The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

#125

Earlier quoted context omitted.

They have tacitly endorsed the leak. https://github.com/facebookresearch/llama/pull/73#issuecomme...

Only because publicly visible actions are worse for them People have gotten DMCA takedown requests from them over Llama repositories

If they were interested in limiting distribution, saying essentially "go ahead and seed this torrent more" is worse for them than doing nothing.

Re: The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

#126

Is this a tactical leak, stemming from a "commoditize your complement" strategy? Open source as a strategic weapon, without having to explain board members/shareholders/whatever that you threw around money on training an open sourced model?

It’s not open source. Llama is proprietary, the license hasn’t changed. Just like the source code to windows leaking doesn’t make windows open source.

Re: The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

#127
post #64

Earlier quoted context omitted.

It's not a bug. It's an architectural defect / limitation in our understanding of how to build AI. That makes it a strictly harder problem that will take longer. And it's not totally clear to me that you'll get there purely with LLMs. LLMs accomplish a good chunk of what we classify as intelligence for sure. But it's missing the cognition / reasoning skills and the open question is whether you can solve that by just…

GPT 4 will admit to not knowing things in many cases where 3.5turbo does not (tested the same prompt), and either will stop there or go off on a "but if it did exist it might go something like this" type continuation. It still hallucinates a lot, but it's not at all clear that this will be all that difficult an issue to solve given the progress.

We generally only hallucinate while dreaming / using our imagination. And we can distinguish those two states. Admitting lack of knowledge is of course good but, for example, if you ask it to write some code that isn’t boilerplate API integrations, it’ll do so happily even when it’s wildly wrong and it can’t tell the difference and that is also the case with GPT4 afaik. Moreover, you can’t solve it through prompt engineering because there’s clearly a lack of context it’s unable to understand to figure out what non trivial thing your asking it.

Re: The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

#128

"The training and serving code, along with an online demo, are publicly available for non-commercial use." (from Vicuna's home page.) In what universe is that "open source"?!

Nothing in the article is open source. A proprietary model got leaked and there are other proprietary apps that are stupidly building on the leaked model.

Re: The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

#129

I've spent an embarassing amount of time since the llamas leaked playing with them, the tools to run them, and writing wrappers for them. They are technically alternatives in the sense that they're incomparably better chat bots than anything in the past. But at least for the 30B and under versions (65B is too big for me to run), no matter what fine tuning is done (alpaca, gpt4all, vicuna, etc), the llamas themselves…

Your conclusion seems not to be warranted since you haven't tried out the 65B model.

I agree, but I think my experience is representative. So far most human people don't have the resources to be able to use 65B. And most small companies / university groups don't have the resources to fine-tune a 65B.

I've talked to a couple dozen people in real time who've played with up to 30B but no one I know has the resources to run the 65B at all or fast enough to actually use and get an opinion of. None of the open source llama projects out there are using 65B in practice (despite support for it) so I think my 30B and under conclusions are applicable to the topic the article covers. I'd love to be wrong and I'm excited for this to change in the future.

Re: The LLama Effect: Leak Sparked a Series of Open Source Alternatives to ChatGPT

#130

Earlier quoted context omitted.

Copyright is a practical right, not an inherent right. The only reasons humans get copyright at all is because it's useful for society to give it to them. The onus should be on OpenAI to prove that it will benefit society overall if AIs are given copyright. We've already decided that many non-human processes/entities don't get copyright because there doesn't seem to be any reason to grant those entities copyright. --…

> The comparison to humans is interesting though, because teaching a human how to do something doesn't grant you copyright over their output. Ehh, in rare cases in can though. If you have someone sign an NDA, they can't go and publish technical details about something confidential that they were trained on. For example, this is fairly common in the tech industry when we send engineers to train on proprietary hardware…

> Ehh, in rare cases in can though. If you have someone sign an NDA, they can't go and publish technical details about something confidential that they were trained on. For example, this is fairly common in the tech industry when we send engineers to train on proprietary hardware or software.

And I think nearly everyone would agree that it would be perfectly fine and reasonable for an AI trained on a proprietary corpus of information to produce copyrightable/secret material in response to questions.

Just because I built an internal corporate search tool, doesn't mean that you get to view its output.

The question at play here is when the AI is trained on information that's in the public commons. The 'teacher' analogy is, in this sense, a very good one.

Post reply on HN