Earlier quoted context omitted.
Yeah, I'm also surprised at how many positive comments are in this thread. I do hate Facebook, but I also love engineers, so I'm not sure how to feel about this one.
> I do hate Facebook, but I also love engineers, so I'm not sure how to feel about this one. "it's complicated". Remember that? :) It's also a great way to avoid many classes of bias. One shouldn't aspire to "feel" in any one way. Embrace the complexity.
Meta Llama 3
881–890 of 965 posts
Re: Meta Llama 3
#882Earlier quoted context omitted.
NVidia, AMD, Microsoft?
Nvidia, maybe. Microsoft, definitely not. Nadella is a successful CEO but is as corporate as they come.
Re: Meta Llama 3
#883Re: Meta Llama 3
#884I just want to express how grateful I am that Zuck and Yann and the rest of the Meta team have adopted an open approach and are sharing the model weights, the tokenizer, information about the training data, etc. They, more than anyone else, are responsible for the explosion of open research and improvement that has happened with things like llama.cpp that now allow you to run quite decent models locally on consumer h…
You can see from Zuck's interviews that he is still an engineer at heart. Every other big tech company has lost that kind of leadership.
Re: Meta Llama 3
#885Earlier quoted context omitted.
Nvidia, maybe. Microsoft, definitely not. Nadella is a successful CEO but is as corporate as they come.
Nadella has such practiced corporate-speak it's impressive. I went to a two-hour talk and Q&A he did, and he didn't communicate a single piece of real information over the whole session. It was entirely HR filler language, the whole time.
Re: Meta Llama 3
#886Earlier quoted context omitted.
> you can never put the genie back in the bottle That's specifically why OpenAI don't release weights, and why everyone who cares about safety talks about laws, and why Yud says the laws only matter if you're willing to enforce them internationally via air strikes. > It’s going to be so satisfying I won't be feeling Schadenfreude if a low budget group or individual takes an open weights model, does a white-box analys…
If some unhinged psycho want to build nerve agents and bombs I think it's laughable to believe an LLM will be the tool that makes a difference in enabling them to do so. As you said the information is already out there - getting info on how to do this stuff is not the barrier you think it is.
Re: Meta Llama 3
#887Earlier quoted context omitted.
"Llama is not open source." This is interesting. Can you point me to an OSI discussion what would constitute an open source license for LLMs? Obviously they have "source" (network definitions) and "training data" and "weights". I'm not aware of any such discussion.
Actually right now the OSI is hosting ongoing discussion this year on what it means for AI to be open source. Here is their latest blog post on the subject: https://opensource.org/blog/open-source-ai-definition-weekly... Here is the latest draft definition: https://hackmd.io/@opensourceinitiative/osaid-0-0-7 And a discussion about the draft: https://discuss.opensource.org/t/draft-v-0-0-7-of-the-open-s...
Re: Meta Llama 3
#8881) It keeps unnecessarily hash-tagging sentences, but listens better than mistral at removing them if I specify
2) It failed my pop culture "Arthas test":
Here's llama3:
>>> who's arthas
A great question!
Arathos, also known as Arathas or Artaxerxes (Persian: اردشیر), was a king of the
Achaemenid Empire, who ruled from 465 to 424 BCE. He was the
son of Xerxes I and Queen Amestris.
Arthas is best known for his military campaigns against
Greece during the Persian Wars.
In particular, he led a successful campaign against
Athens and Eretria in 459-449 BCE, which resulted in the
destruction of these cities
and the imposition of heavy reparations.
However, his later years were marked by a series of
defeats at the hands of the Greeks,
including the Battle of Plataea (479 BCE) and the Battle
of Mycale (479 BCE). This ultimately led to the decline of
the Achaemenid Empire and
its eventual defeat by Alexander the Great in 334 BCE.
So, there you have it! Arthas, a significant figure in
ancient Persian history.
It thought I said "Arathos" yet says "Arthas" at the end lol. Here's mistral: >>> who's arthas
Arthas Menethil is a character from the real-time strategy
game "Warcraft III: Reign of Chaos" and its expansion "The
Frozen Throne," developed
by Blizzard Entertainment. He is also known as the Lich
King.
Arthas was once the prince of Stratholme, in the kingdom
of Lordaeron, and was a promising young knight who was
deeply devoted to his kingdom
and people. However, after a series of tragic events,
including the loss of his mother and the corruption of his
mentor Uther the Lightbringer,
Arthas became consumed by madness and betrayal. He turned
against his own people and eventually became the Lich
King, the leader of the Scourge,
an undead horde that plagued the land of Azeroth.
mistral seems more on-point and has more consistent outputs. but llama3 is better than mistral at following output instructions like "output the same list but dont use numbers" (mistral insists on either bullets or numbers, have to say "in naturally flowing sentences" to get it to work). This might mean llama3 is better to build with (esp. RAG), even if it's less knowledgeable.It's not gpt-4 level that's for sure. I think mistral is still closest to gpt-4 that I've tried.
Re: Meta Llama 3
#889Earlier quoted context omitted.
I haven't tried Llama 3 yet, but Llama 2 is indeed extremely "safe." (I'm old enough to remember when AI safety was about not having AI take over the world and kill all humans, not when it might offend a Puritan's sexual sensibilities or hurt somebody's feelings, so I hate using the word "safe" for it, but I can't think of a better word that others would understand). It's not quite as bad as Gemini, but in the same c…
So whereabouts are you that a "Puritan's sexual sensibilities" holds any sway?
Re: Meta Llama 3
#890Earlier quoted context omitted.
1. Free rlhf 2. They cookie the hell out of you to breadcrumb your journey around the web. They don't need you to login to get what they need, much like Google
Do they really need “free RLHF”? As I understand it, RLHF needs relatively little data to work and its quality matters - I would expect paid and trained labellers to do a much better job than Joey Keyboard clicking past a “which helped you more” prompt whilst trying to generate an email.
Modern captchas are self driving object labelers; you just need a few to "agree" to know what the right answer is.