Live data from Hacker News

Meta Llama 3

llama.meta.com

711–720 of 965 posts

Re: Meta Llama 3

#711
post #643

Earlier quoted context omitted.

I don’t understand how the idea of open source become some sort of pseudo-legalistic purity test on everything. Models aren’t code, some of the concepts of open source code don’t map 1:1 to freely available models. In spirit I think this is “open source”, and I think that’s how the majority of people think. Turning everything into some sort of theological debate takes away a lot of credit that Meta deserves. Google i…

> Turning everything into some sort of theological debate takes away a lot of credit that Meta deserves. It's not theological, it's the misuse of a specific legal definition that we all have interest in maintaining. "Freely available models" or "open license" are accurate. Other companies keeping things for themselves doesn't warp reality, or the existing definitions we use to describe it. Giving them the credit they…

Hate to break it to you but there’s a thousand court cases a day precisely because “specific legal definition” is a surprisingly flexible concept depending on context. Likewise when new technologies emerge it often requires reappraisal and interpretation of existing laws, even if that reappraisal is simply extending the old law to the new context.

Re: Meta Llama 3

#712

https://github.com/meta-llama/llama3/blob/main/LICENSE Llama is not open source. It's corporate freeware with some generous allowances. Open source licenses are a well defined thing. Meta marketing saying otherwise doesn't mean they get to usurp the meaning of a well understood and commonly used understanding of the term "open source." https://opensource.org/license Nothing about Meta's license is open source. It's a…

"Llama is not open source." This is interesting. Can you point me to an OSI discussion what would constitute an open source license for LLMs? Obviously they have "source" (network definitions) and "training data" and "weights". I'm not aware of any such discussion.

Actually right now the OSI is hosting ongoing discussion this year on what it means for AI to be open source. Here is their latest blog post on the subject:

https://opensource.org/blog/open-source-ai-definition-weekly...

Here is the latest draft definition:

https://hackmd.io/@opensourceinitiative/osaid-0-0-7

And a discussion about the draft:

https://discuss.opensource.org/t/draft-v-0-0-7-of-the-open-s...

Re: Meta Llama 3

#713

How do they plan to make money with this? They can even make money with their 24K GPU cluster as IaaS if they want to. Even Google is gatekeeping its best Gemini model behind. https://web.archive.org/web/20240000000000*/https://filebin.... https://web.archive.org/web/20240419035112/https://s3.filebi...

Are those links connected to your comment?

Re: Meta Llama 3

#714
post #542

Earlier quoted context omitted.

They are aiming to beat the current GPT4 and stand a fair chance, they are unlikly to hold the crown for long.

Right because the very little I've heard out of Sam Altman this year hinting at future updates suggests that there's something coming before we turn our calendars to 2025. So equaling or mildly exceeding GPT-4 will certainly be welcome, but could amount to a temporary stint as king of the mountain.

This is always the case.

But the fact that open models are beating state of the art from 6 months ago is really telling just how little moat there is around AI.

Re: Meta Llama 3

#715

https://github.com/meta-llama/llama3/blob/main/LICENSE Llama is not open source. It's corporate freeware with some generous allowances. Open source licenses are a well defined thing. Meta marketing saying otherwise doesn't mean they get to usurp the meaning of a well understood and commonly used understanding of the term "open source." https://opensource.org/license Nothing about Meta's license is open source. It's a…

(We detached this subthread from https://news.ycombinator.com/item?id=40077832)

Re: Meta Llama 3

#717

Earlier quoted context omitted.

The day it's released, Llama-3-405B will be running on someone's Mac Studio. These models aren't that big. It'll be fine, just like Llama-2.

Maybe at 1 or 2 bits of quantization! Even the Macs with the most unified RAM are maxxed out with much smaller models than 405b (especially since it's a dense model and not a MOE).

You can build a $6,000 machine with 12 channels DDR5 memory that's big enough to hold an 8bit quantized model. The generation speed is abysmal of course.

Anything better than that starts at 200k per machine and goes up from there.

Not something you can run at home, but definitely within the budget of most medium sized firms to buy one.

Re: Meta Llama 3

#718
Meta Llama 3 8B vs Claude Haiku according to their press releases if anyone else was curious

| Benchmark | Llama3 8B | Claude Haiku |

| ------------- | ----------- | ------------ |

| MMLU ____ | 68.4 ____ | 75.2 _______ |

| GPQA ____ | 34.2 ____ | 33.3 _______ |

| HumanEval | 62.2 ____ | 75.9 _______ |

| GSM-8K __ | 79.6 ____ | 88.9 _______ |

| MATH ____ | 30.0 ____ | 40.9 _______ |

Re: Meta Llama 3

#719

Meta Llama 3 8B vs Claude Haiku according to their press releases if anyone else was curious | Benchmark | Llama3 8B | Claude Haiku | | ------------- | ----------- | ------------ | | MMLU ____ | 68.4 ____ | 75.2 _______ | | GPQA ____ | 34.2 ____ | 33.3 _______ | | HumanEval | 62.2 ____ | 75.9 _______ | | GSM-8K __ | 79.6 ____ | 88.9 _______ | | MATH ____ | 30.0 ____ | 40.9 _______ |

This llama model some made it run on an iphone. https://x.com/1littlecoder/status/1781076849335861637?s=46

Re: Meta Llama 3

#720
I just saw an ad on Facebook for a Meta AI image generator. The ad featured a little girl doing prompt engineering, then being excited at the picture of the unicorn it made. It made me sad :(
Post reply on HN