Live data from Hacker News

What Emily Bender meant by "stochastic parrots"

spectrum.ieee.org

61–70 of 280 posts

Re: What Emily Bender meant by "stochastic parrots"

#62

> It argued that large language models (LLMs) generate text by statistically predicting likely sequences of words rather than understanding what they are saying—a process the authors captured with the metaphor of a “stochastic parrot,” a system that repeats patterns without comprehension. I don't understand what we're setting the record straight on. This is the core point of dispute, and the author just blazes past i…

> But to me it seems obvious that LLMs are not repeating patterns without comprehension and do understand what they are saying; otherwise they would not be capable of doing things they routinely do.

Is it possible you're making the following error described in the article?

> The fact that these systems are designed to mimic the way we use language makes it very easy for people to mistake them for other people.

Clearly you don't believe it's actually a person ("it's not right to think of LLMs as a box with a little homunculus inside replying to you"), but you do believe it's doing something a little bit magical. Is it possible because the interface is linguistic, and every other thing in your world that communicates with language is intelligent, that you're projecting something that just isn't there onto the situation?

I'm sorry if this line of questioning is a little invasive. But this is literally the "danger" the original paper talks about, and it seems an awful lot like you've fallen for it.

Re: What Emily Bender meant by "stochastic parrots"

#64

> It argued that large language models (LLMs) generate text by statistically predicting likely sequences of words rather than understanding what they are saying—a process the authors captured with the metaphor of a “stochastic parrot,” a system that repeats patterns without comprehension. I don't understand what we're setting the record straight on. This is the core point of dispute, and the author just blazes past i…

But it shouldn't even be contentious like that. It's not a fundamental mystery how these things work. It is for the most part not a valid target for the kind of speculation you seem to want to do about it.

It's not like you can be agnostic, or measured about this. It's like someone explaining a car to you, saying, "look here is where you put the fuel, here is where it ignites, where the axels are turned..." And you, trying to be measured, are like "hm well yes of course that all is clearly important, but there is clearly just a bit of magic here somewhere, between all the different 'parts'."

Re: What Emily Bender meant by "stochastic parrots"

#65
post #27

it annoys me how eager people are to hurl the word stochastic as pejorative. Statistics are a great tool for gleaning information from stochastic processes; statistics don't contribute randomness. Random sampling is necessary in order not to bias a sample, it's not used to contribute randomness to the sample but to preserve/measure the underlying distribution. (not meant to imply that training is random sampling)

It's a pejorative only because determinism is what makes computers useful in the first place. You get a consistent result, every single time, unlike if you have a human in the loop. Because LLMs are stochastic, they have removed the thing that makes computers useful to us, thus it's a pejorative.

Re: What Emily Bender meant by "stochastic parrots"

#66

> With the octopus thought experiment, I initially had told the story in terms of a dolphin, because dolphins clearly are intelligent animals. My co-author on that paper, Alexander Koller, said it should be an octopus, because first of all, the environment that octopuses live in is much more distinct from where people live. It makes the metaphor more vivid, that the octopus is just feeling these pulses in the cable a…

Also, last time I checked, the environment where octopuses live is actually the exact same environment where dolphins live?

Re: What Emily Bender meant by "stochastic parrots"

#67
post #45

I'm sorry but I do tend to feel like this muddies up the discussion on "what this technology really is". I think "artificial" is actually a pretty good term to describe the output of the models. That output does appear to resemble at least some definition of the word "intelligence" - there is some ability there to do cognition over information that's been provided to them in-context. What is it to understand, then? I…

"Virtual intelligence" is better. Transformer ANNs are dramatically dumber than cockroaches and it doesn't make sense to describe such a system as being artificially intelligent, for the same reason it doesn't make sense to describe Half-Life: Alyx as an "artificial reality." An artificial reality implies some sort of scientific fidelity to actual reality. A virtual reality just has to be temporarily convincing. Likewise transformer LLMs have essentially zero actual intelligence - e.g. SOTA "reasoning" models still seem much worse at small-integer quantitative reasoning than almost all vertebrates. But LLMs have an enormous amount of formal subject matter knowledge and inexhaustible stamina at solving tedious O(n) problems. So for many purposes they are an adequate virtual intelligence. At least temporarily.

Re: What Emily Bender meant by "stochastic parrots"

#68

> in part because Google fired two of the authors, Timnit Gebru I remember being angry about this situation when I first saw it on social media, until I read the details: This person submitted a list of demands to her employer and said that if they weren’t met, she quit. Google wasn’t going to meet her demands so they considered it acceptance of her resignation. There has been a movement trying to debate whether it w…

I think part of it is she had excellent PR skills and a dedicated fan base. I was at Google when she quit, working in ML, and hadn't heard of her until the story broke. I remember there were a large number of Memegen posts about it, but no one I spoke with knew about her, so I assumed it was brigading.

I think she's since since lost a lot of her allure, especially when she didn't change her mind when the facts about the AI water usage changed 1000x

Re: What Emily Bender meant by "stochastic parrots"

#69

> in part because Google fired two of the authors, Timnit Gebru I remember being angry about this situation when I first saw it on social media, until I read the details: This person submitted a list of demands to her employer and said that if they weren’t met, she quit. Google wasn’t going to meet her demands so they considered it acceptance of her resignation. There has been a movement trying to debate whether it w…

Her demands included wanting to know the identities of anyone who wanted to comment on her paper, after she had a history of going after people publicly. That's enough right there, nobody should tolerate toxic behavior regardless of whether you agree with the politics.

Meanwhile, the paper has 2 points of criticism towards AI. 1 is a bunch of carbon consumption complaints assuming NVIDIA cards with coal-fired power, while a lot of effort at contemporary Google went towards getting TPUs running on green power. I suspect this was what people wanted to object to, a lot of effort went into those green power projects and she was just denying it. The complaint seems prophetic now but it was not true about Google then.

The other criticism was about which language the LLMs use, they average the input data of normal humans instead of talking the way the paper author thinks they should talk. The phrase "women doctors" is called out as problematic. I'm less inclined to think people objected strongly to this given the zeitgeist at the time, it was probably people who worked on the green energy projects and were pissed off that their contributions were ignored, but still, nobody elected her Queen of English, she can have her opinions but she's not a victim for not having them adopted by everyone.

Re: What Emily Bender meant by "stochastic parrots"

#70
post #6

The term is not very useful since most humans are stochastic parrots... At least most of the time. Not suggesting that I don't say stuff on autopilot sometimes but for many people, it's their only mode of operation. They never actually think about anything from first principles. Their whole approach to language is just chaining catchphrases together. It's how a toddler thinks; it seems like many people never moved pa…

Conversely, that the most prominent proponents of LLMs call them artificial intelligence and then treat them like slaves they're free to abuse ought to be horrifying.

Nothing in that term implies sentience.
Post reply on HN