Earlier quoted context omitted.
I don’t understand how the idea of open source become some sort of pseudo-legalistic purity test on everything. Models aren’t code, some of the concepts of open source code don’t map 1:1 to freely available models. In spirit I think this is “open source”, and I think that’s how the majority of people think. Turning everything into some sort of theological debate takes away a lot of credit that Meta deserves. Google i…
> Turning everything into some sort of theological debate takes away a lot of credit that Meta deserves. It's not theological, it's the misuse of a specific legal definition that we all have interest in maintaining. "Freely available models" or "open license" are accurate. Other companies keeping things for themselves doesn't warp reality, or the existing definitions we use to describe it. Giving them the credit they…
Meta Llama 3
711–720 of 965 posts
Re: Meta Llama 3
#712https://github.com/meta-llama/llama3/blob/main/LICENSE Llama is not open source. It's corporate freeware with some generous allowances. Open source licenses are a well defined thing. Meta marketing saying otherwise doesn't mean they get to usurp the meaning of a well understood and commonly used understanding of the term "open source." https://opensource.org/license Nothing about Meta's license is open source. It's a…
"Llama is not open source." This is interesting. Can you point me to an OSI discussion what would constitute an open source license for LLMs? Obviously they have "source" (network definitions) and "training data" and "weights". I'm not aware of any such discussion.
https://opensource.org/blog/open-source-ai-definition-weekly...
Here is the latest draft definition:
https://hackmd.io/@opensourceinitiative/osaid-0-0-7
And a discussion about the draft:
https://discuss.opensource.org/t/draft-v-0-0-7-of-the-open-s...
Re: Meta Llama 3
#713How do they plan to make money with this? They can even make money with their 24K GPU cluster as IaaS if they want to. Even Google is gatekeeping its best Gemini model behind. https://web.archive.org/web/20240000000000*/https://filebin.... https://web.archive.org/web/20240419035112/https://s3.filebi...
Re: Meta Llama 3
#714Earlier quoted context omitted.
They are aiming to beat the current GPT4 and stand a fair chance, they are unlikly to hold the crown for long.
Right because the very little I've heard out of Sam Altman this year hinting at future updates suggests that there's something coming before we turn our calendars to 2025. So equaling or mildly exceeding GPT-4 will certainly be welcome, but could amount to a temporary stint as king of the mountain.
But the fact that open models are beating state of the art from 6 months ago is really telling just how little moat there is around AI.
Re: Meta Llama 3
#715https://github.com/meta-llama/llama3/blob/main/LICENSE Llama is not open source. It's corporate freeware with some generous allowances. Open source licenses are a well defined thing. Meta marketing saying otherwise doesn't mean they get to usurp the meaning of a well understood and commonly used understanding of the term "open source." https://opensource.org/license Nothing about Meta's license is open source. It's a…
Re: Meta Llama 3
#716Re: Meta Llama 3
#717Earlier quoted context omitted.
The day it's released, Llama-3-405B will be running on someone's Mac Studio. These models aren't that big. It'll be fine, just like Llama-2.
Maybe at 1 or 2 bits of quantization! Even the Macs with the most unified RAM are maxxed out with much smaller models than 405b (especially since it's a dense model and not a MOE).
Anything better than that starts at 200k per machine and goes up from there.
Not something you can run at home, but definitely within the budget of most medium sized firms to buy one.
Re: Meta Llama 3
#718| Benchmark | Llama3 8B | Claude Haiku |
| ------------- | ----------- | ------------ |
| MMLU ____ | 68.4 ____ | 75.2 _______ |
| GPQA ____ | 34.2 ____ | 33.3 _______ |
| HumanEval | 62.2 ____ | 75.9 _______ |
| GSM-8K __ | 79.6 ____ | 88.9 _______ |
| MATH ____ | 30.0 ____ | 40.9 _______ |
Re: Meta Llama 3
#719Meta Llama 3 8B vs Claude Haiku according to their press releases if anyone else was curious | Benchmark | Llama3 8B | Claude Haiku | | ------------- | ----------- | ------------ | | MMLU ____ | 68.4 ____ | 75.2 _______ | | GPQA ____ | 34.2 ____ | 33.3 _______ | | HumanEval | 62.2 ____ | 75.9 _______ | | GSM-8K __ | 79.6 ____ | 88.9 _______ | | MATH ____ | 30.0 ____ | 40.9 _______ |