Live data from Hacker News

Meta Llama 3

llama.meta.com

881–890 of 965 posts

Re: Meta Llama 3

#881
post #673

Earlier quoted context omitted.

Yeah, I'm also surprised at how many positive comments are in this thread. I do hate Facebook, but I also love engineers, so I'm not sure how to feel about this one.

> I do hate Facebook, but I also love engineers, so I'm not sure how to feel about this one. "it's complicated". Remember that? :) It's also a great way to avoid many classes of bias. One shouldn't aspire to "feel" in any one way. Embrace the complexity.

You're right. It's just, of course, easier to feel one extreme or the other.

Re: Meta Llama 3

#882
post #532

Earlier quoted context omitted.

NVidia, AMD, Microsoft?

Nvidia, maybe. Microsoft, definitely not. Nadella is a successful CEO but is as corporate as they come.

Nadella has such practiced corporate-speak it's impressive. I went to a two-hour talk and Q&A he did, and he didn't communicate a single piece of real information over the whole session. It was entirely HR filler language, the whole time.

Re: Meta Llama 3

#883
When I make a request, Meta begins to answer it (I can see the answer appear) and almost immediately, a negative response shows up indicating they’re working on it (ex: I ask if it’s capable of working in French, Meta indicates that it can, the message disappears and is replaced by “I don’t understand French yet, but I’m working on it. I will send you a message when we can talk in French”). What gives?

Re: Meta Llama 3

#884
post #240

I just want to express how grateful I am that Zuck and Yann and the rest of the Meta team have adopted an open approach and are sharing the model weights, the tokenizer, information about the training data, etc. They, more than anyone else, are responsible for the explosion of open research and improvement that has happened with things like llama.cpp that now allow you to run quite decent models locally on consumer h…

You can see from Zuck's interviews that he is still an engineer at heart. Every other big tech company has lost that kind of leadership.

This is both their biggest strength and weakness

Re: Meta Llama 3

#885
post #532

Earlier quoted context omitted.

Nvidia, maybe. Microsoft, definitely not. Nadella is a successful CEO but is as corporate as they come.

Nadella has such practiced corporate-speak it's impressive. I went to a two-hour talk and Q&A he did, and he didn't communicate a single piece of real information over the whole session. It was entirely HR filler language, the whole time.

[deleted]

Re: Meta Llama 3

#886
post #757

Earlier quoted context omitted.

> you can never put the genie back in the bottle That's specifically why OpenAI don't release weights, and why everyone who cares about safety talks about laws, and why Yud says the laws only matter if you're willing to enforce them internationally via air strikes. > It’s going to be so satisfying I won't be feeling Schadenfreude if a low budget group or individual takes an open weights model, does a white-box analys…

If some unhinged psycho want to build nerve agents and bombs I think it's laughable to believe an LLM will be the tool that makes a difference in enabling them to do so. As you said the information is already out there - getting info on how to do this stuff is not the barrier you think it is.

when all you want is to hurt then every tool looks like a weapon.

Re: Meta Llama 3

#887

Earlier quoted context omitted.

"Llama is not open source." This is interesting. Can you point me to an OSI discussion what would constitute an open source license for LLMs? Obviously they have "source" (network definitions) and "training data" and "weights". I'm not aware of any such discussion.

Actually right now the OSI is hosting ongoing discussion this year on what it means for AI to be open source. Here is their latest blog post on the subject: https://opensource.org/blog/open-source-ai-definition-weekly... Here is the latest draft definition: https://hackmd.io/@opensourceinitiative/osaid-0-0-7 And a discussion about the draft: https://discuss.opensource.org/t/draft-v-0-0-7-of-the-open-s...

Thanks!

Re: Meta Llama 3

#888
First experience with llama3 (via ollama)

1) It keeps unnecessarily hash-tagging sentences, but listens better than mistral at removing them if I specify

2) It failed my pop culture "Arthas test":

Here's llama3:

  >>> who's arthas
  
  A great question!

  Arathos, also known as Arathas or Artaxerxes (Persian: اردشیر), was a king of the 
  Achaemenid Empire, who ruled from 465 to 424 BCE. He was the 
  son of Xerxes I and Queen Amestris.

  Arthas is best known for his military campaigns against 
  Greece during the Persian Wars. 
  In particular, he led a successful campaign against 
  Athens and Eretria in 459-449 BCE, which resulted in the 
  destruction of these cities 
  and the imposition of heavy reparations.

  However, his later years were marked by a series of 
  defeats at the hands of the Greeks, 
  including the Battle of Plataea (479 BCE) and the Battle
  of Mycale (479 BCE). This ultimately led to the decline of 
  the Achaemenid Empire and 
  its eventual defeat by Alexander the Great in 334 BCE.

  So, there you have it! Arthas, a significant figure in 
  ancient Persian history.
It thought I said "Arathos" yet says "Arthas" at the end lol. Here's mistral:

  >>> who's arthas

  Arthas Menethil is a character from the real-time strategy 
  game "Warcraft III: Reign of Chaos" and its expansion "The 
  Frozen Throne," developed
  by Blizzard Entertainment. He is also known as the Lich 
  King.

  Arthas was once the prince of Stratholme, in the kingdom 
  of Lordaeron, and was a promising young knight who was 
  deeply devoted to his kingdom 
  and people. However, after a series of tragic events, 
  including the loss of his mother and the corruption of his 
  mentor Uther the Lightbringer, 
  Arthas became consumed by madness and betrayal. He turned 
  against his own people and eventually became the Lich 
  King, the leader of the Scourge,
  an undead horde that plagued the land of Azeroth.
mistral seems more on-point and has more consistent outputs. but llama3 is better than mistral at following output instructions like "output the same list but dont use numbers" (mistral insists on either bullets or numbers, have to say "in naturally flowing sentences" to get it to work). This might mean llama3 is better to build with (esp. RAG), even if it's less knowledgeable.

It's not gpt-4 level that's for sure. I think mistral is still closest to gpt-4 that I've tried.

Re: Meta Llama 3

#889

Earlier quoted context omitted.

I haven't tried Llama 3 yet, but Llama 2 is indeed extremely "safe." (I'm old enough to remember when AI safety was about not having AI take over the world and kill all humans, not when it might offend a Puritan's sexual sensibilities or hurt somebody's feelings, so I hate using the word "safe" for it, but I can't think of a better word that others would understand). It's not quite as bad as Gemini, but in the same c…

So whereabouts are you that a "Puritan's sexual sensibilities" holds any sway?

It’s everywhere. The entire USA has been devolving into New Puritan nonsense in many ways since the sexual revolution… which is bizarre.

Re: Meta Llama 3

#890
post #732

Earlier quoted context omitted.

1. Free rlhf 2. They cookie the hell out of you to breadcrumb your journey around the web. They don't need you to login to get what they need, much like Google

Do they really need “free RLHF”? As I understand it, RLHF needs relatively little data to work and its quality matters - I would expect paid and trained labellers to do a much better job than Joey Keyboard clicking past a “which helped you more” prompt whilst trying to generate an email.

Absolutely.

Modern captchas are self driving object labelers; you just need a few to "agree" to know what the right answer is.

Post reply on HN