Live data from Hacker News

Llama 2

ai.meta.com

331–340 of 860 posts

Re: Llama 2

#331

Earlier quoted context omitted.

That's just a trade. If we assume "charity" is "altruism," then by definition there must be no benefit to the giver.

How can it be a trade if one party gave nothing to the other party? If one company gets good PR and a group gets something for free, how is that a trade? One party can benefit and give nothing, while the other party still benefits. I've literally never done anything charitable by your definition then, because i do it because it makes me feel good. I like helping others. Perhaps the only charitable companies or people…

Ask yourself, would your charity exist without your benefits? If no than you've always done it for your self interest.

Re: Llama 2

#332
post #275

Hey HN, we've released tools that make it easy to test LLaMa 2 and add it to your own app! Model playground here: https://llama2.ai Hosted chat API here: https://replicate.com/a16z-infra/llama13b-v2-chat If you want to just play with the model, llama2.ai is a very easy way to do it. So far, we’ve found the performance is similar to GPT-3.5 with far fewer parameters, especially for creative tasks and interactions. Dev…

Still fails my hippo test! > Yes, hippos are excellent swimmers. They spend most of their time in the water, where they feed on aquatic plants and escape the heat of the savannah. In fact, hippos are one of the best swimmers among all land mammals. But that's fine. Most do. Hippos don't swim. They walk or hop/skip at best underwater.

There's a few prompts that I use with every model to compare them. One of the simplest ones is:

> When does the bowl of the winds get used in the wheel of time books?

LLaMA2 fails pretty hard:

> The Bowl of the Winds is a significant artifact in the Wheel of Time series by Robert Jordan. It is first introduced in the third book, "The Dragon Reborn," and plays a crucial role in the series throughout the rest of the books. The Bowl of the Wines is a powerful tool that can control the winds and is used by the Aes Sedai to travel long distances and to escape danger. It is used by the male Aes Sedai to channel the True Power and to perform various feats of magic.

For what it's worth Bard is the only model that I've seen get this question correct with most others hallucinating terrible answers. I'm not sure what it is about this question that trips LLMs up so much but they produce notably bad results when prompted with it.

> Please write a function in JavaScript that takes in a string as input and returns true if it contains a valid roman numeral and false otherwise.

Is another test that I like, which so far no LLM I've tested passes but GPT-4 comes very close.

Here LLaMA2 also fails pretty hard, though I thought this follow up response was pretty funny:

> The function would return true for 'IIIIII' because it contains the Roman numeral 'IV'.

Re: Llama 2

#333

Key detail from release: > If, on the Llama 2 version release date, the monthly active users of the products or services made available by or for Licensee, or Licensee’s affiliates, is greater than 700 million monthly active users in the preceding calendar month, you must request a license from Meta, which Meta may grant to you in its sole discretion, and you are not authorized to exercise any of the rights under thi…

Should have been an asterisk on the headline like “free … for commercial* use”

Re: Llama 2

#334

Hey HN, we've released tools that make it easy to test LLaMa 2 and add it to your own app! Model playground here: https://llama2.ai Hosted chat API here: https://replicate.com/a16z-infra/llama13b-v2-chat If you want to just play with the model, llama2.ai is a very easy way to do it. So far, we’ve found the performance is similar to GPT-3.5 with far fewer parameters, especially for creative tasks and interactions. Dev…

>If you want to just play with the model, llama2.ai is a very easy way to do it.

Currently suffering from a hug of death

Re: Llama 2

#335

Earlier quoted context omitted.

That's just a trade. If we assume "charity" is "altruism," then by definition there must be no benefit to the giver.

I don't think that's even possible, but if it was it would be a disaster because humans don't work that way. We respond to incentive. When giving to charity, the incentive can be as simple as "I feel good" but it's still an incentive.

Some do what's right even if it doesn't feel good. The best charity can be painful.

Re: Llama 2

#336
post #275

Hey HN, we've released tools that make it easy to test LLaMa 2 and add it to your own app! Model playground here: https://llama2.ai Hosted chat API here: https://replicate.com/a16z-infra/llama13b-v2-chat If you want to just play with the model, llama2.ai is a very easy way to do it. So far, we’ve found the performance is similar to GPT-3.5 with far fewer parameters, especially for creative tasks and interactions. Dev…

Still fails my hippo test! > Yes, hippos are excellent swimmers. They spend most of their time in the water, where they feed on aquatic plants and escape the heat of the savannah. In fact, hippos are one of the best swimmers among all land mammals. But that's fine. Most do. Hippos don't swim. They walk or hop/skip at best underwater.

This is a pedantic non issue and has nothing to do with the overall thread.

Re: Llama 2

#337
post #304
post #275

Earlier quoted context omitted.

Still fails my hippo test! > Yes, hippos are excellent swimmers. They spend most of their time in the water, where they feed on aquatic plants and escape the heat of the savannah. In fact, hippos are one of the best swimmers among all land mammals. But that's fine. Most do. Hippos don't swim. They walk or hop/skip at best underwater.

This test seems to be testing the ability of it to accurately convey fine details about the world. If that's what you're looking for it's a useful test, but if you're looking for a language model and not a general knowledge model I'm not sure it's super relevant. The average person probably couldn't tell you if a hippo swims either, or having been informed about how a hippo locomotes whether or not that counts as swi…

So it's more designed for a superficial chat?

Re: Llama 2

#338
post #320
post #305

Earlier quoted context omitted.

You're just being overly pedantic. They hold their breath, fully submerge, control their buoyancy, and propel themselves through water. Also known as swimming.

Nah, this is often not considered swimming in major publications and by zoos. National Geographic https://www.nationalgeographic.com/animals/mammals/facts/hip... > Hippos cannot swim or breathe underwater, and unlike most mammals they are so dense that they cannot float. Instead, they walk or run along the bottom of the riverbed. Because their eyes and nostrils are located on the top of their heads, they can still se…

> among particularly good mammal swimmers

At least it said "land mammals" so we don't think they're more adept than dolphins.

Re: Llama 2

#339

Key detail from release: > If, on the Llama 2 version release date, the monthly active users of the products or services made available by or for Licensee, or Licensee’s affiliates, is greater than 700 million monthly active users in the preceding calendar month, you must request a license from Meta, which Meta may grant to you in its sole discretion, and you are not authorized to exercise any of the rights under thi…

https://blogs.microsoft.com/blog/2023/07/18/microsoft-and-me... I think this is effectively an Apple + Amazon + Google ban? (MS employee, just noticing interesting intersection of announcements and licensing).

Interesting, so Meta doesn't want to pay for the hardware and they partner with MS to use Azure. On the other hand, MS provides hardware for free, hoping they consolidate their investment in AI.

Re: Llama 2

#340
post #60

Earlier quoted context omitted.

Google has far better models than llama based models. They just simply don't put them facing the public. It is pretty ridiculous that they essentially just set a marketing team with no programming experience to write Bard, but that shouldn't fool anyone into believing they don't have capable models in Google. If Deepmind were to actually provide what they have in some usable form, it would likely be quite good. Despi…

I’ve been hearing “Google has secret better models” for 7 months now. Maybe some UFOs in the hangers at Moffett Field too?

Do you realize that LLaMA-1 is just a very slightly smaller, comparably performing replication of Chinchilla [1], which DeepMind had completed a year prior to LLaMA's release? And has RLHF-ed into a suitable chatbot "Sparrow" [2] months earlier than ChatGPT was launched?

To assume that Google doesn't have anything competitive with Meta is to say that their papers just so happen to contain recipes for Meta's models but they've arrived at those not through training and benchmarking but by divination and bullshitting. This, let us say, does not sound plausible.

Then again, Microsoft uses LLaMA for research, and they should theoretically have some ability to get stuff from OpenAI. Evidently this isn't how any of this works, huh.

1. https://arxiv.org/abs/2203.15556

2. https://en.wikipedia.org/wiki/Sparrow_(bot)

Post reply on HN