Live data from Hacker News

Why Meta’s latest large language model survived only three days online

technologyreview.com

81–90 of 126 posts

Re: Why Meta’s latest large language model survived only three days online

#81

This outcome from using a large language model to mimic reasoning isn’t surprising. What’s surprising is Yan LeCun’s childish and petty reaction to this entirely foreseeable series of events: > Galactica demo is off line for now. It’s no longer possible to have some fun by casually misusing it. Happy? He’s supposedly an expert in this sort of thing

He is an expert. Much better than you’ll ever be.

So did researcher in NLP became better or worse off after demo taken down and why?

Re: Why Meta’s latest large language model survived only three days online

#82
post #25

Earlier quoted context omitted.

Weird, it shows up as 1:40 long for me. It's the last 5 minutes of the episode, where they claim GPT-3 is an all knowing machine that will generate factual responses to any question in a way that's superior to google search.

My dad is a doctor who oversees residents. Seems like half the time they call him for advice he just puts their question into gpt-3 and regurgitates it’s answer, so bill isn’t the only one.

If that's actually happening--and I am both skeptical and terrified that it is--it seems awfully close to malpractice or even (criminal) negligence.

Re: Why Meta’s latest large language model survived only three days online

#84
post #71

Earlier quoted context omitted.

The world's most expensive Lorem Ipsum generator?

I think it's a search engine with a bad curation/ranking algorithm. It's trained with a corpus of research papers it mines from in response to a search prompt. It's a bit like if Google were to haphazardly compose a website from the first 20 pages of search results, or worse. Composition is the novelity here, and we should judge it based on how well it can select and compose. Turns out not that well yet; judgement is…

A startup called Cuil (https://en.wikipedia.org/wiki/Cuil) tried exactly the strategy you suggest in jest: synthesize articles by mashing up search results. It was a disaster, and widely mocked for how easy it was to get Cuil to produce absolute nonsense from straightforward prompts. When your starting point is "untrustworthy nonsense", it is an uphill battle in both technology and PR to arrive at "trustworthy synthesis", if it is indeed possible at all.

Re: Why Meta’s latest large language model survived only three days online

#85
post #34
post #16

Earlier quoted context omitted.

I would urge him to put it back online, it is interesting and can be useful. Just don't make a press release about it, journalists ruin everything.

The problem is not journalist, it's about how Meta and LeCun presented it. They presented it as "you should trust what it says and use to write papers", then hid in the small lines "oh actually really don't do that". You can't have your cake and eat it AND complain about being called out on it.

> "you should trust what it says and use to write papers"

They really said something like that?

Re: Why Meta’s latest large language model survived only three days online

#86

"A fundamental problem with Galactica is that it is not able to distinguish truth from falsehood," In true science, it is exceptionally hard to distinguish truth from falsehood for many of the interesting subjects. It can take decades of work to reach consensus on what is "truth." Physics in the early 20th century is a great example of this debate.

Exactly. Also why identifying "misinformation" is a fool's errand, since yesterday's misinformation is today's truth.

> Also why identifying "misinformation" is a fool's errand

Seems easy enough: as long as the content is inoffensive and fits into the Overton Window then it's not misinformation.

Re: Why Meta’s latest large language model survived only three days online

#87
post #63

Earlier quoted context omitted.

This wasnt meant to generate valid scientific papers, and Lecun said so too. It generates interesting associations. It rambles sometimes and goes on tangents that are sometimes relevant sometimes not. It can inform you of related ideas that you were not aware of. It's like a fuzzy google scholar. It is in no way valid publishable research, but it's like a bicycle for researchers. At least that was what i managed to f…

As far as I understand (and reading their Limitations page also), the system is quite likely to simply invent facts, particularly in niche fields - which may well mislead you and lead on a wild goose chase.

If you're a) believing the output wholesale or b) not occasionally going down random rabbit holes to explore new ideas (even ones that seemed silly or wasteful on the surface) you're probably doing something wrong as it is.

Re: Why Meta’s latest large language model survived only three days online

#88
post #3

It’s algorithmically/randomly generating text without understanding. What it the proper way of using it? Fake papers? Political bs? Bad Hemingway (or Shakespeare or Chaucer or…). It’s noise that looks like sentences.

I think the point was actually to demo self supervised learning techniques (which is LeCun's schtick) in a way that was a bit flashy and accessible to the public. Fun, easily shareable on social media, generates some buzz about FB AI, etc.

Clearly pitching it as an actual, authoritative source of info was not the right call

Re: Why Meta’s latest large language model survived only three days online

#89

This software is excellent for pseudo science. For example, young earth peddlers will be able to generate entire mambo jambo references and use them to indoctrinate more people.

Are you sure this is an actual, real life problem?

I'm still waiting for all of the FUD the GPT3 doomers were warning us would happen. It's been out for a year now.

Either our existing reputation systems are pretty resilient or no one has yet seen any actual value in generating generic text at scale for malicious purposes.

Re: Why Meta’s latest large language model survived only three days online

#90
I think an always correct version of Galactica can't be ML-only based. In the end, every "fact" goes back to the question "what are truthful facts?". What we read on Wikipedia? What scientist claim? What the majority of humanity thinks?

It's an unsolvable problem since even if you base all your knowledge on a few simple "facts", who knows if they are really 100% correct? E.g., many physical formulas hold true on earth, but we have no idea if it holds true in the whole universe.

Post reply on HN