Live data from Hacker News

Llama 2

ai.meta.com

141–150 of 860 posts

Re: Llama 2

#141

Hey HN, we've released tools that make it easy to test LLaMa 2 and add it to your own app! Model playground here: https://llama2.ai Hosted chat API here: https://replicate.com/a16z-infra/llama13b-v2-chat If you want to just play with the model, llama2.ai is a very easy way to do it. So far, we’ve found the performance is similar to GPT-3.5 with far fewer parameters, especially for creative tasks and interactions. Dev…

Will Llama 2 also work as a drop-in in existing tools like llama.cpp, or does it require different / updated tools?

Re: Llama 2

#142
post #19

Earlier quoted context omitted.

I think more Apple. It's not like Google or Microsoft would want to use LLaMA when they have fully capable models themselves. I wouldn't be surprised if Amazon does as well. Apple is the big laggard in terms of big tech and complex neural network models.

Apple would absolutely not want to use a competitors, or any other, public LLM. They want to own the whole stack, and will want to have their own secret source as part of it. It's not like they don't have the capital to invest in training...

[deleted]

Re: Llama 2

#143

Earlier quoted context omitted.

Apple would absolutely not want to use a competitors, or any other, public LLM. They want to own the whole stack, and will want to have their own secret source as part of it. It's not like they don't have the capital to invest in training...

Apple does not have the capability to train a LLM currently.

I very much doubt that.

Re: Llama 2

#144

Earlier quoted context omitted.

> Bard release, it would have made more sense for them to have a more limited release of a better model for PR reasons than what actually happened. Yes I would agree with you if Google wasn't set on to full on panic mode by their investors about releasing something vs Open AI due to Chat GPT's buzz. Bard was just a "hey we can do this too" thing, it was released half assed, had next to no marketing or hype. Vertex AI…

I can already tell you that PaLM is not anywhere near as good and PaLM-2 is at least not as good before RLHF. Not going to keep replying, believe what you want about Google's capabilities

ok now I am confused, as Meta themselves say Palm-2 is better than Llama 2?

> Llama 2 70B results are on par or better than PaLM (540B) (Chowdhery et al., 2022) on almost all benchmarks. There is still a large gap in performance between Llama 2 70B and GPT-4 and PaLM-2-L.

https://scontent.fsyd7-1.fna.fbcdn.net/v/t39.2365-6/10000000...

If Google's publically available model is better Llama 2 already then why is it so inconceivable that they'd have private models that are better than their public ones which are better than LLama already.

Palm-2 isn't better than GPT-4 but the convo was about better than Llama models no?

Re: Llama 2

#145
post #30

Earlier quoted context omitted.

> OpenAI's ChatGPT hit 100 million MAUs in January, and has gone down since. poor reading of the numbers. one guy at a bank pulled up similarweb and guesstimated 100m registered users and it went viral. whisper numbers were closer to 50m. but in the 6 months since they have certainly crossed 100m and probably are north of 500m, and only recently dipped.

How do you find Whisper numbers, it’s open source yea?

Whisper numbers are numbers that are secretly shared among industry insiders, not the usage numbers of OpenAI's Whisper.

Re: Llama 2

#146

Hey HN, we've released tools that make it easy to test LLaMa 2 and add it to your own app! Model playground here: https://llama2.ai Hosted chat API here: https://replicate.com/a16z-infra/llama13b-v2-chat If you want to just play with the model, llama2.ai is a very easy way to do it. So far, we’ve found the performance is similar to GPT-3.5 with far fewer parameters, especially for creative tasks and interactions. Dev…

Seeing a16z w/early access, enough to build multiple tools in advance, is a very unpleasant reminder of insularity and self-dealing of SV elites. My greatest hope for AI is no one falls for this kind of stuff the way we did for mobile.

And yet here we are a few weeks after that with a free to use model that cost millions to develop and is open to everyone.

I think you’re taking an unwarranted entitled view.

Re: Llama 2

#147

Earlier quoted context omitted.

You have to agree to any terms they might think of in the future. Clicking download, they claim you agree to their privacy policy which they claim they can update on a whim Google's privacy policy, for example, was updated stealthfully to let them claim rights over every piece of IP you post on the internet that their crawlers can get to

> Google's privacy policy, for example, lets them claim rights over every piece of IP you post on the internet without protecting it behind a paywall This is a nonsense. They added a disclaimer basically warning that LLMs might learn some of your personal data from the public web, because that’s part of the training data. A privacy policy is not a contract that you agree to, it’s just a notice of where/when your data…

Google it. They're just laundering it through their ai first

Re: Llama 2

#148

Hey HN, we've released tools that make it easy to test LLaMa 2 and add it to your own app! Model playground here: https://llama2.ai Hosted chat API here: https://replicate.com/a16z-infra/llama13b-v2-chat If you want to just play with the model, llama2.ai is a very easy way to do it. So far, we’ve found the performance is similar to GPT-3.5 with far fewer parameters, especially for creative tasks and interactions. Dev…

How does one apply for a job with the the internal A16Z teams experimenting with this?

Re: Llama 2

#149

Hey HN, we've released tools that make it easy to test LLaMa 2 and add it to your own app! Model playground here: https://llama2.ai Hosted chat API here: https://replicate.com/a16z-infra/llama13b-v2-chat If you want to just play with the model, llama2.ai is a very easy way to do it. So far, we’ve found the performance is similar to GPT-3.5 with far fewer parameters, especially for creative tasks and interactions. Dev…

Seeing a16z w/early access, enough to build multiple tools in advance, is a very unpleasant reminder of insularity and self-dealing of SV elites. My greatest hope for AI is no one falls for this kind of stuff the way we did for mobile.

e: Oh - this is a16z, so yeah probably early access - scratch my additional comments

I agree that I don't like early/insider stuff

That said - I believe Llama 2 is architecturally identical to the previous one and given that they are using 13B it is probably just a drag and drop bin replacement and reload your servers.

We all knew Llama 2 was coming so it might be within the capabilities of a hungry startup with no early access.

Re: Llama 2

#150

Earlier quoted context omitted.

"In addition to open-source models, we also compare Llama 2 70B results to closed-source models. As shown in Table 4, Llama 2 70B is close to GPT-3.5 (OpenAI, 2023) on MMLU and GSM8K, but there is a significant gap on coding benchmarks. Llama 2 70B results are on par or better than PaLM (540B) (Chowdhery et al., 2022) on almost all benchmarks. There is still a large gap in performance between Llama 2 70B and GPT-4 an…

it's not open source

This quote does not talk about Llama being open source.
Post reply on HN