Live data from Hacker News

Meta Llama 3

llama.meta.com

681–690 of 965 posts

Re: Meta Llama 3

#681

I just want to express how grateful I am that Zuck and Yann and the rest of the Meta team have adopted an open approach and are sharing the model weights, the tokenizer, information about the training data, etc. They, more than anyone else, are responsible for the explosion of open research and improvement that has happened with things like llama.cpp that now allow you to run quite decent models locally on consumer h…

> but also to not use pessimistic AI "doomerism" as an excuse to hide the crown jewels and put it behind a centralized API with a gatekeeper because of "AI safety risks."

AI safety risk is substantial. It is also testable. (There are prediction markets on it, for example.) Of course, some companies may latch onto various valid arguments for insincere reasons.

I'd challenge everyone to closely compare ideas such as "open source software is better" versus "state of the art trained AI models are better developed in the open". The exact same arguments do NOT work for both.

It is one thing to publish papers about e.g. transformers. It is another thing to publish the weights of something like GPT 3.5+; it might theoretically be a matter of degree, but that matter of degree makes a real difference, if only in terms of time. Time matters because it gives people and society some time to respond.

Software security reports are often made privately or embargoed. Why? We want to give people and companies time to defend their systems.

Now consider this thought-experiment: assume LLMs (and their hybrid derivatives) enable perhaps 1,000,000 new kinds of cyberattacks, 1,000 new bioweapon attacks, and so on. Are there are a correspondingly large number of defensive benefits? This is the crux of the question I think. First, I don't expect we're going to get a good assessment of the overall "balance". Second, any claims of "balance" are beside the point, because these attacks and defenses don't simply cancel each other out. The distribution of the AI-fueled capability advance will probably ratchet up risk and instability.

Open source software's benefits stem from the assumption that bugs get shallower with more eyes. More eyes means that the open source product gets stronger defensively.

With LLMs that publish their weights, both the research and the implementations is out; you can't get guardrails. The closest analogue to an "OSS security report" would take the form of "I just got your LLM to design a novel biological weapon. Do you think you can use it to design an antidote?"

A systematic risk-averse person might want to ask: what happens if we enumerate all offensive vs defensive technological shifts? Should we reasonably believe that the benefits outweigh the risks?

Unfortunately, the companies making these decisions aren't bearing the risks. This huge externality both pisses me off and scares the shit out of me.

Re: Meta Llama 3

#682
post #309

Earlier quoted context omitted.

They're commoditizing their complement [0][1], inasmuch as LLMs are a complement of social media and advertising (which I think they are). They've made it harder for competitors like Google or TikTok to compete with Meta on the basis of "we have a super secret proprietary AI that no one else has that's leagues better than anything else". If everyone has access to a high quality AI (perhaps not the world's best, but c…

Yes. And, could potentially diminish OpenAI/MS. Once everyone can do it, then OpenAI value would evaporate.

Very similar to Tesla and EVs

Re: Meta Llama 3

#683

I just want to express how grateful I am that Zuck and Yann and the rest of the Meta team have adopted an open approach and are sharing the model weights, the tokenizer, information about the training data, etc. They, more than anyone else, are responsible for the explosion of open research and improvement that has happened with things like llama.cpp that now allow you to run quite decent models locally on consumer h…

It's like Elon saying: we have open sourced our patents, use them. Well, use the old patents and stay behind forever....

Exactly.

Re: Meta Llama 3

#684

I just want to express how grateful I am that Zuck and Yann and the rest of the Meta team have adopted an open approach and are sharing the model weights, the tokenizer, information about the training data, etc. They, more than anyone else, are responsible for the explosion of open research and improvement that has happened with things like llama.cpp that now allow you to run quite decent models locally on consumer h…

The quickest way to disabuse yourself of this notion is to login to Facebook. You’ll remember that Zuck makes money from the scummiest pool of trash and misinformation the world has ever seen. He’s basically the Web 2.0 tabloid newspaper king. I don’t really care how much the AI team open sources, the world would be a better place if the entire company ceased to exist.

Yeah lmao, people are giving meta way too much credit here tbh.

Re: Meta Llama 3

#685
post #680

Earlier quoted context omitted.

It certainly doesn't, I'm running the 7B locally with ollama It provided a lot more detail about the case, but does not have current information. It hallucinated the question about juror count, or maybe confused it with a different case seems more likely, one of the E Jean Carroll cases or the SDNY Trump Org financial fraud case?

You: how many jurists have been selected in the Trump trial in New York? Meta AI: A full jury of 12 people has been selected for former President Donald Trump's trial in New York City, in addition to one alternate ¹. The selection process will continue in order to select five more alternates, though it is hoped that the selection process will be finished tomorrow ². Once all alternates have been selected, opening sta…

Yup, the Meta hosted system is much more than LLaMA 3. Seems to have RAG, search, and/or tool usage

Re: Meta Llama 3

#688
post #512
post #497

Earlier quoted context omitted.

Let's be honest that he's probably not doing it due to goodness of his heart. He's most likely trying to commoditize the models so he can sell their complement. It's a strategy Joel Spolsky had talked about in the past (for those of you who remember who that is). I'm not sure what the complement of AI models is that Meta can sell exactly, so maybe it's not a good strategy but I'm certain it's a strategy of some sort

Also keep in mind that it's still a proprietary model. Meta gets all the benefits of open source contributions and testing while retaining exclusive business use.

Very wrong.

Llama is usable by any company under 700M MAU.

Re: Meta Llama 3

#689

I just want to express how grateful I am that Zuck and Yann and the rest of the Meta team have adopted an open approach and are sharing the model weights, the tokenizer, information about the training data, etc. They, more than anyone else, are responsible for the explosion of open research and improvement that has happened with things like llama.cpp that now allow you to run quite decent models locally on consumer h…

The quickest way to disabuse yourself of this notion is to login to Facebook. You’ll remember that Zuck makes money from the scummiest pool of trash and misinformation the world has ever seen. He’s basically the Web 2.0 tabloid newspaper king. I don’t really care how much the AI team open sources, the world would be a better place if the entire company ceased to exist.

[flagged]
Post reply on HN