Live data from Hacker News

"Hallucinating" AIs sound creative, but let's not celebrate being wrong

thereader.mitpress.mit.edu

31–40 of 196 posts

Re: "Hallucinating" AIs sound creative, but let's not celebrate being wrong

#31

I never liked that the term "hallucinating" ended up sticking. The AI doesn't know something so it just invents something. We usually call that "bullshitting", or in more polite crowds, "lying".

I agree, at the very least use the psychologically accurate word for making shit up: confabulating

Re: "Hallucinating" AIs sound creative, but let's not celebrate being wrong

#32
Contra the headling, we should "celebrate being wrong" – when it's wrong in fast, interesting, & voluminous ways that can then be filtered or corrected.

That's how lots of science, innovation, & learning work: generate many superficially-plausible candidates via a fast-and-loose process, then refine with a more rigorous evaluation.

That AIs, in the form of LLMs, are now doing this so well was unexpected, and progress in checking 'hallucinations' is proceeding very fast.

(Fortunately, the article is less dismissive than the headline, recognizing these model's potential & mainly urging an understanding of the limitations.)

Re: "Hallucinating" AIs sound creative, but let's not celebrate being wrong

#33

I never liked that the term "hallucinating" ended up sticking. The AI doesn't know something so it just invents something. We usually call that "bullshitting", or in more polite crowds, "lying".

Lying implies intent... bullshitting may actually a somewhat better term than hallucinating, in the sense that the hallucinator doesn't direct their hallucination so as to please an audience, while a bullshitter often does.

Re: "Hallucinating" AIs sound creative, but let's not celebrate being wrong

#34

This is an odd article. To me, it seems like the "creative" arts are an ideal arena for AI. After all, there's no such thing as "wrong" art. The article says "well, sometimes what it makes is bad". Well big deal. A lot of human-created art is awful too.

> This is an odd article. To me, it seems like the "creative" arts are an ideal arena for AI. After all, there's no such thing as "wrong" art.

Yeah but who wants to consume art purely generated by AI (that is, not human-created with AI support)? Most art sites have had blanket bans, or at least required tagging, on ai-generated art because people hate it so much.

Or to put it another way: why are you in the comment section of Hacker News, and not just asking ChatGPT to generate social media comments on the article?

Re: "Hallucinating" AIs sound creative, but let's not celebrate being wrong

#35

Curiously enough, I don't get the same result as the author on plain ChatGPT-4. https://chat.openai.com/share/e213e0bd-2838-45e2-9942-e52954... https://chat.openai.com/share/64bc62d3-042c-40c4-8d50-8e28ce... Hilariously, plugging the example in the article into Bing enhanced ChatGPT-4, ChatGPT-4 w/ Bing hallucinates because of that very article! https://chat.openai.com/share/4fe50933-8436-44ad-a778-6297ca... If you t…

> Hilariously, plugging the example in the article into Bing enhanced ChatGPT-4, ChatGPT-4 w/ Bing hallucinates because of that very article!

Bing does not use the actual GPT 4 model. It's almost certainly a lower parameter model (like 180 billion vs > 1 trillion), or at the very least heavily quantized. That's why it makes more mistakes in your tests.

Re: "Hallucinating" AIs sound creative, but let's not celebrate being wrong

#36

Key sentence for me: > It might be better to say that everything GPT does is a hallucination , since a state of non-hallucination, of checking the validity of something against some external perception, is absent from these models. I try to explain this to people who are obsessed with using ChatGPT to tell them things. So far I've been telling them something like: "it does not attempt to provide you valid information…

> "it does not attempt to provide you valid information, it's optimizing for what would read like a reasonable continuation of the conversation, which is really not the same thing."

No, but it's quite close - because "reasonable" is positively correlated with "valid, correct information".

Re: "Hallucinating" AIs sound creative, but let's not celebrate being wrong

#37
Think of requiring a minimum amount of hallucination in AI output as a safety mechanism, like mixing the distinctive odor into a propane tank. The odor is a signal that there's a gas leak that must be dealt with. A high minimum amount of hallucinations is a signal that the source is untrustworthy and must be checked. Hallucinations may turn out to be a feature that protects against over reliance. And defers the need for a Butlerian jihad.

Re: "Hallucinating" AIs sound creative, but let's not celebrate being wrong

#39
post #32

Contra the headling, we should "celebrate being wrong" – when it's wrong in fast, interesting, & voluminous ways that can then be filtered or corrected. That's how lots of science, innovation, & learning work: generate many superficially-plausible candidates via a fast-and-loose process, then refine with a more rigorous evaluation. That AIs, in the form of LLMs, are now doing this so well was unexpected, and progress…

Science progresses because people are testing plausible ideas not random ideas.

We don’t do human trials on random drugs, we do human trials on promising ones. Animal trials are more open but even then candidates are carefully considered as being viable. Things are even more open the earlier you are in drug discovery, but we aren’t testing molecules using Dubnium or most elements on the periodic table.

Similarly Ecology isn’t studying what happens when you introduce each type of salt water fish into lakes because the general assumption is they would just die thus saving you from preforming millions of experiments. That basic check for plausibility is extraordinarily valuable across all sciences.

There’s no sacred cows here. You do need to validate plausibility just like anything else, but you don’t need to do an exhaustive search across every possibility.

Post reply on HN