Earlier quoted context omitted.
What does a submarine do? Submarine? I suppose you "drive" a submarine which is getting to the idea: submarines don't swim because ultimately they are "driven"? I guess the issue is we don't make up a new word for what submarines do, we just don't use human words. I think the above poster gets a little distracted by suggesting the models are creative which itself is disputed. Perhaps a better term, like above, would…
I really like that, I think it has the right amount of distance. They don't write, they model writing. We're very used to "all models are wrong, some are useful", "the map is not the territory", etc.
A non-anthropomorphized view of LLMs
241–250 of 432 posts
Re: A non-anthropomorphized view of LLMs
#242Re: A non-anthropomorphized view of LLMs
#243Earlier quoted context omitted.
Where exactly in my description do I invoke consciousness? Where does the description given imply that consciousness is required in any way? The fact that there's a non-obvious emergent phenomena which is apparently responsible for your subjective experience, and that it's possible to provide a superficially accurate description of you as a system without referencing that phenomena in any way, is my entire point. The…
My bad, we are saying the same thing. I misinterpreted your last sentence as saying this simplistic view of the brain you described does not account for consciousness.
Re: A non-anthropomorphized view of LLMs
#244Earlier quoted context omitted.
On the contrary, anthropomorphism IMO is the main problem with narratives around LLMs - people are genuinely talking about them thinking and reasoning when they are doing nothing of that sort (actively encouraged by the companies selling them) and it is completely distorting discussions on their use and perceptions of their utility.
When I see these debates it's always the other way around - one person speaks colloquially about an LLM's behavior, and then somebody else jumps on them for supposedly believing the model is conscious, just because the speaker said "the model thinks.." or "the model knows.." or whatever. To be honest the impression I've gotten is that some people are just very interested in talking about not anthropomorphizing AI, an…
Re: A non-anthropomorphized view of LLMs
#245Earlier quoted context omitted.
On the contrary, anthropomorphism IMO is the main problem with narratives around LLMs - people are genuinely talking about them thinking and reasoning when they are doing nothing of that sort (actively encouraged by the companies selling them) and it is completely distorting discussions on their use and perceptions of their utility.
I kinda agree with both of you. It might be a required abstraction, but it's a leaky one. Long before LLMs, I would talk about classes / functions / modules like "it then does this, decides the epsilon is too low, chops it up and adds it to the list". The difference I guess it was only to a technical crowd and nobody would mistake this for anything it wasn't. Everybody know that "it" didn't "decide" anything. With AI…
This made me think, when will we see LLMs do the same; rereading what they just sent, and editing and correcting their output again :P
Re: A non-anthropomorphized view of LLMs
#246Earlier quoted context omitted.
"predirence" -> prediction meets inference and it sounds a bit like preference
Except -ence is a regular morph, and you would rather suffix it to predict(at)-. And prediction is already an hyponym of inference. Why not just use inference then?
Inference would be the part that is deliberately learned and drawn from conclusions based on the training set, like in the "classic" sense of statistical learning.
Re: A non-anthropomorphized view of LLMs
#247Earlier quoted context omitted.
I'll say it once more: I think it is useful to distinguish between autoregressive and recurrent architectures. A clear way to make that distinction is to agree that the recurrent architecture has hidden state, while the autoregressive one does not. A recurrent model has some point in a space that "encapsulates its understanding". This space is "hidden" in the sense that it doesn't correspond to text tokens or any oth…
I'll also point out what is most important part from your original message: > LLMs have hidden state not necessarily directly reflected in the tokens being produced, and it is possible for LLMs to output tokens in opposition to this hidden state to achieve longer-term outcomes (or predictions, if you prefer). But what does it mean for an LLM to output a token in opposition to its hidden state? If there's a longer-ter…
Re: A non-anthropomorphized view of LLMs
#248Earlier quoted context omitted.
I kinda agree with both of you. It might be a required abstraction, but it's a leaky one. Long before LLMs, I would talk about classes / functions / modules like "it then does this, decides the epsilon is too low, chops it up and adds it to the list". The difference I guess it was only to a technical crowd and nobody would mistake this for anything it wasn't. Everybody know that "it" didn't "decide" anything. With AI…
I mean you can boil anything down to it's building blocks and make it seem like it didn't 'decide' anything. When you as a human decide something, your brain and it's neurons just made some connections with an output signal sent to other parts that resulting in your body 'doing' something. I don't think LLMs are sentient or any bullshit like that, but I do think people are too quick to write them off before really th…
These are very different and knowledge is not intelligence.
Re: A non-anthropomorphized view of LLMs
#249Earlier quoted context omitted.
The "point" of not anthropomorphizing is to refrain from judgement until a more solid abstraction appears. The problem with explaining LLMs in terms of human behaviour is that, while we don't clearly understand what the LLM is doing, we understand human cognition even less! There is literally no predictive power in the abstraction "The LLM is thinking like I am thinking". It gives you no mechanism to evaluate what ta…
> Why don't LLMs get frustrated with you if you ask them the same question repeatedly? To be fair, I have had a strong sense of Gemini in particular becoming a lot more frustrated with me than GPT or Claude. Yesterday I had it ensuring me that it was doing a great job, it was just me not understanding the challenge but it would break it down step by step just to make it obvious to me (only to repeat the same errors,…
Re: A non-anthropomorphized view of LLMs
#250Here's a question for you: how do you reconcile that these stochastic mapping are starting to realize and comment on the fact that tests are being performed on them when processing data?