Live data from Hacker News

Meta Llama 3

llama.meta.com

721–730 of 965 posts

Re: Meta Llama 3

#721
post #590

Earlier quoted context omitted.

I'm based on LLaMA 2, which is a type of transformer language model developed by Meta AI. LLaMA 2 is a more advanced version of the original LLaMA model, with improved performance and capabilities. I'm a specific instance of LLaMA 2, trained on a massive dataset of text from the internet, books, and other sources, and fine-tuned for conversational AI applications. My knowledge cutoff is December 2022, and I'm constan…

Strange. The Llama 3 model card mentions that the knowledge cutoff dates are March 2023 for the 8B version and December 2023 for the 70B version ( https://github.com/meta-llama/llama3/blob/main/MODEL_CARD.md )

Maybe a typo?

Re: Meta Llama 3

#722
post #4

They've got a console for it as well, https://www.meta.ai/ And announcing a lot of integration across the Meta product suite, https://about.fb.com/news/2024/04/meta-ai-assistant-built-wi... Neglected to include comparisons against GPT-4-Turbo or Claude Opus, so I guess it's far from being a frontier model. We'll see how it fares in the LLM Arena.

Losers & Winners from Llama-3-400B Matching 'Claude 3 Opus' etc.. Losers: - Nvidia Stock : lid on GPU growth in the coming year or two as "Nation states" use Llama-3/Llama-4 instead spending $$$ on GPU for own models, same goes with big corporations. - OpenAI & Sam: hard to raise speculated $100 Billion, Given GPT-4/GPT-5 advances are visible now. - Google : diminished AI superiority posture Winners: - AMD, intel: th…

Disagree on Nvidia, most folks fine-tune model. Proof: there are about 20k models in huggingface derived from llama 2, all of them trained on Nvidia GPUs.

Re: Meta Llama 3

#723
post #4

They've got a console for it as well, https://www.meta.ai/ And announcing a lot of integration across the Meta product suite, https://about.fb.com/news/2024/04/meta-ai-assistant-built-wi... Neglected to include comparisons against GPT-4-Turbo or Claude Opus, so I guess it's far from being a frontier model. We'll see how it fares in the LLM Arena.

Tried a few queries and was surprised how fast it responded vs how slow chatgpt can be. Responses seemed just as good too.

Inference speed is not a great metric given the horizontal scalability of LLMs.

Re: Meta Llama 3

#724
post #443

Earlier quoted context omitted.

>We’re rolling out Meta AI in English in more than a dozen countries outside of the US. Now, people will have access to Meta AI in Australia, Canada, Ghana, Jamaica, Malawi, New Zealand, Nigeria, Pakistan, Singapore, South Africa, Uganda, Zambia and Zimbabwe — and we’re just getting started. https://about.fb.com/news/2024/04/meta-ai-assistant-built-wi...

That's a strange list of nations, isn't it? I wonder what their logic is.

GPU server locations, maybe?

Re: Meta Llama 3

#725

I just want to express how grateful I am that Zuck and Yann and the rest of the Meta team have adopted an open approach and are sharing the model weights, the tokenizer, information about the training data, etc. They, more than anyone else, are responsible for the explosion of open research and improvement that has happened with things like llama.cpp that now allow you to run quite decent models locally on consumer h…

I actually think Mr Zuckerburg is maturing and has a chance of developing a public persona of being decent person!

I say public persona, as I've never met him, and have no idea what he is like as a person on an individual level.

Maturing in general and studying martial arts is likely to be a contributing factor.

Re: Meta Llama 3

#726

"You’ll also soon be able to test multimodal Meta AI on our Ray-Ban Meta smart glasses." Now this is interesting. I've been thinking for some time now that traditional computer/smartphone interfaces are on the way out for all but a few niche applications. Instead, everyone will have their own AI assistant, which you'll interact with naturally the same way as you interact with other people. Need something visual? Just…

There are a dozen different services to get the last X days of MSFT stock price. If you’re interested in stocks, you probably have a favorite already. Why would someone need an AI assistant for this?

[deleted]

Re: Meta Llama 3

#727
post #240

I just want to express how grateful I am that Zuck and Yann and the rest of the Meta team have adopted an open approach and are sharing the model weights, the tokenizer, information about the training data, etc. They, more than anyone else, are responsible for the explosion of open research and improvement that has happened with things like llama.cpp that now allow you to run quite decent models locally on consumer h…

You can see from Zuck's interviews that he is still an engineer at heart. Every other big tech company has lost that kind of leadership.

Apple being the most egregious example IMHO.

Purely my opinion as a long time Apple fan, but I cant help but think that Tim Cook's polices are harming the Apple brand in ways that we wont see for a few years.

Much like Balmer did at Microsoft.

But who knows - I'm just making conversation :-)

Re: Meta Llama 3

#728

Earlier quoted context omitted.

They didn't compare against the best models because they were trying to do "in class" comparisons, and the 70B model is in the same class as Sonnet (which they do compare against) and GPT3.5 (which is much worse than sonnet). If they're beating sonnet that means they're going to be within stabbing distance of opus and gpt4 for most tasks, with the only major difference probably arising in extremely difficult reasonin…

Llama is open weight, not open source. They don’t release all the things you need to reproduce their weights.

Has anyone tested how close you need to be to the weights for copyright purposes?

Re: Meta Llama 3

#729

Earlier quoted context omitted.

Right because the very little I've heard out of Sam Altman this year hinting at future updates suggests that there's something coming before we turn our calendars to 2025. So equaling or mildly exceeding GPT-4 will certainly be welcome, but could amount to a temporary stint as king of the mountain.

This is always the case. But the fact that open models are beating state of the art from 6 months ago is really telling just how little moat there is around AI.

Unless you are NVidia.

Re: Meta Llama 3

#730
post #522

Earlier quoted context omitted.

Seems like a year or two of MMA has done way more for his charisma than whatever media training he's done over the years. He's a lot more natural in interviews now.

Alternatively, he’s completely relaxed here because he knows what he’s doing is genuinely good and people will support it. That’s gotta be a lot less stressful than, say, a senate hearing.

The net positive outcome of AI is still to evaluate, same with social media and he still pays by selling our data.
Post reply on HN