Live data from Hacker News

Meta Llama 3

llama.meta.com

831–840 of 965 posts

Re: Meta Llama 3

#831
post #754

Earlier quoted context omitted.

So whereabouts are you that a "Puritan's sexual sensibilities" holds any sway?

I think the point is Silicon Valley is such a place. "Visible nipples? The defining characteristic of all mammals, which infants necessarily have to put in their mouths to feed? On this website? Your account has been banned!" Meanwhile in Berlin, topless calendars in shopping malls and spinning-cube billboards for Dildo King all over the place.

TO be fair, its one of the most popular terms people search...

So, lets not pretend its something that isnt arousing.

Re: Meta Llama 3

#832
post #229

Earlier quoted context omitted.

Very interesting part around 5 mins in where Zuck says that they bought a shit ton of H100 GPUs a few years ago to build the recommendation engine for Reels to compete with TikTok (2x what they needed at the time, just to be safe), and now they are accidentally one of the very few companies out there with enough GPU capacity to train LLMs at this scale.

The only thing the Reels algorithm is showing me are videos of ladies with fat butts. Now, I must admit, I may have clicked once on such a video. Should I now be damned to spend an eternity in ass hell?

It’s easy to populate your feed with things you specifically want to watch: watch the stuff you’re interested in and swipe on the things that don’t interest you.

Re: Meta Llama 3

#833
post #761

Earlier quoted context omitted.

so run it locally, local version is not guarded

My locally-hosted llama3 actually craps itself if I ask it to answer in other languages. It's pretty hilarious. Has been working flawlessly (and impressively fast) for everything in English, then does hilarious glitches in other languages. Eg right now to show it here, I say "Write me a poem about a digital pirate in Danish": Digitalen Pirat På nettet sejler han, En digital pirat, fri og farlig. Han har øjnene på de…

The training data is 95% English, foreign language is not going to be its strongest strength.

Re: Meta Llama 3

#834

Earlier quoted context omitted.

> I just want to express how grateful I am that Zuck Praise for him at HN? It should be enough of a reason for him to pop a champagne today

Yeah, I'm also surprised at how many positive comments are in this thread. I do hate Facebook, but I also love engineers, so I'm not sure how to feel about this one.

One of the many perks of releasing open-ish models, React, and many other widely used tools over the years. Meta might be the big tech whose open source projects are most widely used. That gives you some dev goodwill, even though your main products profit from some pretty bad stuff.

Re: Meta Llama 3

#835

Earlier quoted context omitted.

You can build a machine that can run 70b models at great TpS speeds for around 30-60k. That same machine could almost certainly run a 400b model with "useable" speeds. Obviously much slower than current ChatGPT speeds but still, that kind of machine is well within the means of wealthy hobbyists/highly compensated SWEs and small firms.

I just tested llama3:70b with ollama on my old AMD ThreadRipper Pro 3965WX workstation (16-core Zen4 with 8 DDR4 mem channels), with a single RTX 4090. Got 3.5-4 tokens/s, GPU compute was <20% busy (~90W) and the 16 CPU cores / 32 threads were about 50% busy.

Jesus that's the old one?

Re: Meta Llama 3

#836
post #755

Earlier quoted context omitted.

For sure. I just started watching the new Dwarkesh interview with Zuck that was just released ( https://t.co/f4h7ko0M7q ) and you can just tell from the first few minutes that he simply has a different level of enthusiasm and passion and level of engagement than 99% of big tech CEOs.

I've never heard of this person, but many of the questions he asks Zuck show a total lack of any insight in this field. How did this interview even happen?

He’s built up an impressive amount of clout over a short period of time, mostly by interviewing interesting guests on his podcast while not boring listeners to death (unlike a certain other interviewer with high-caliber guests that shall remain nameless).

Re: Meta Llama 3

#837
post #836
post #755

Earlier quoted context omitted.

I've never heard of this person, but many of the questions he asks Zuck show a total lack of any insight in this field. How did this interview even happen?

He’s built up an impressive amount of clout over a short period of time, mostly by interviewing interesting guests on his podcast while not boring listeners to death (unlike a certain other interviewer with high-caliber guests that shall remain nameless).

What's the meaning of life though, and why is it love?

Re: Meta Llama 3

#838

Earlier quoted context omitted.

Right because the very little I've heard out of Sam Altman this year hinting at future updates suggests that there's something coming before we turn our calendars to 2025. So equaling or mildly exceeding GPT-4 will certainly be welcome, but could amount to a temporary stint as king of the mountain.

This is always the case. But the fact that open models are beating state of the art from 6 months ago is really telling just how little moat there is around AI.

>This is always the case.

I mean anyone can throw out self evident general truisms about how there will always be new models and always new top dogs. It's a good generic assumption but I feel like I can make generic assumptions and general truisms just as well as the next person.

I'm more interested in divining in specific terms who we consider to be at the top currently, tomorrow and the day after tomorrow based on the specific things that have been reported thus far. And interestingly, thus far, the process hasn't been one of a regular rotation of temporary top dogs. It's been one top dog, Open AI's GPT, I would say that it currently is still, and when looking at what the future holds, it appears that it may have a temporary interruption before it once again is the top dog, so to speak.

That's not to say it'll always be the case but it seems like that's what our near future timeline has in store based on reporting, and it's piecing that near future together that I'm most interested in.

Re: Meta Llama 3

#839

Earlier quoted context omitted.

You lead with a command to be honest and then immediately speculate on private unknowable motivations and then attribute, without evidence, his decision to a strategy you can't describe. What is this? Someone said something nice, and you need to "restore balance"

They said something naive, not just "nice". It's good to correct the naivete. For example, as we speak, Zuck is lobbying congress to ban Tiktok. Putting aside whether you think it should be banned, this is clearly a cynical strategy with pure self interest in mind. He's trying to monopolize. Whatever Zuck's strategy with open source is, it's just a strategy. Much like AMD is pursuing that strategy. They're corporatio…

What was said that was naive?

Re: Meta Llama 3

#840

I tried generating a Chinese rap song, and it did generate a pretty good rap. However, upon completion, it deleted the response, and showed > I don’t understand Chinese yet, but I’m working on it. I will send you a message when we can talk in Chinese. I tried some other languages and the same. It will generate non-English language, but once its done, the response is deleted and replaced with the message

Crazy that this bug is still happening 12hrs later
Post reply on HN