I have small kids, toddlers, who can already speak the language but still developing their "sense of the world" or "theory of mind" if you will. Maybe it's just me, but talking to toddlers often reminds me of interacting with LLMs, where you would have this realization from time to time "oh, they don't get this, need to break down more to explain". Of course LLM has more elaborate language skills due to its exposure…
> Maybe it's just me, but talking to toddlers often reminds me of interacting with LLMs It's not just you. It hit me almost a year ago, when I realized my then 3.5yo daughter has a noticeable context window of about 30 seconds - whenever she went on her random rant/story, anything she didn't repeat within 30 seconds would permanently fall out of the story and never be mentioned again. It also made me realize why smal…
Large language models lack deep insights or a theory of mind
211–220 of 270 posts
Re: Large language models lack deep insights or a theory of mind
#212Re: Large language models lack deep insights or a theory of mind
#213Earlier quoted context omitted.
Are you saying every thought you’ve ever had is a logical consequence of all prior thoughts with a definite traceable lineage? (Assuming your mind was too dumb as some point and a thought emerged that originated all future thoughts)
I'm saying that I'm extremely confident that... 1. Most people (including Buddhists) think that their internal monologue comprises their thoughts, in majority or even in total. 2. Can't conceive of the possibility of a thought existing other than expressed in their spoken language 3. Find it difficult or impossible to suppress their inner monologue. 4. When successful at suppressing it believe that their thoughts hav…
It doesn't need to be, and isn't, a conscious decision.
Re: Large language models lack deep insights or a theory of mind
#214Earlier quoted context omitted.
The idea is indeed logical and stupidly obvious, once you learn the basics of what "optimization" means, or what "recursion" is. > I find that hard to believe. Ever watched Terminator? Terminator has fuck all to do with recursive self-improvement. Don't confuse people who grew up on sci-fi with people who casually went to see Terminator or some other pop-culture artifact featuring some kind of "AI". > OK, then. You'r…
>>> I've never actually read that in fiction. >> I find that hard to believe. Ever watched Terminator? > Terminator has fuck all to do with recursive self-improvement. Don't confuse people who grew up on sci-fi with people who casually went to see Terminator or some other pop-culture artifact featuring some kind of "AI". You're not following the thread. The future timeline in Terminator does involve something like an…
Yes. That's distinctly different from Skynet iterating on itself a billion times to make itself smarter, which AFAIK (I'm not up to date with full Terminatorverse, but then, most people aren't either), isn't something that happened in that story.
> The popularity of that and similar sci-fi makes that claim that someone has never encountered it hard to believe.
Again, there's very little in mass-market sci-fi of what we're discussing here. And most people, including many in tech, have a hard time wrapping their heads around the idea of a feedback loop, so no, I don't think it's something readily available from mass-market sci-fi.
(But the more niche, better thought-out works, will teach you feedback loops, and this is just one of the ways recursive self-improvement becomes an obvious idea.)
> So how has that been going? Those things should probably be labeled "science fiction."
Eugenics? We had to ban it and create such a strong cultural (and legal) repulsive field around it, that it impedes biotech and medical research.
Designer babies? Weren't attempts made in China recently? And in the West, we're already correcting congenital defects, so all in all, it's less "science fiction", and more "science someone is going to apply soon, if they haven't already".
> How do you know it would have an easier job optimizing itself than humans have?
Because it was created by us, using processes and media that are strongly optimized for malleability. Software, algorithms, digital data, optimization models. All well-defined (and comprehensible to an AGI, by definition) - unlike our own minds, which were not made by us but by a dumb, random process, and that the brains are made of stupidly complex nanotech instead of simple transistors is not helping.
Also because the kind of models we're now worried about gain capability through an optimization process that's open-ended, and limited only by availability of training data and compute. So if e.g. a successor of GPT-4 were to become AGI, it would be set up for recursive self-improvement from day one.
> How do you know there isn't some fundamental contradiction in the concept of "superintelligence" that these fantasies are based on? Or even just some practical resource limits that makes the fantasy impossible?
Maybe, but what makes you think this is the case? We know of some fundamental limits to compute, but we're very, very far from hitting them. Otherwise, I don't know of anything that would put a cap on intelligence at around human level. Remember: by the very nature of evolution, we're the dumbest possible beings capable of learning and building a technological civilization. There may be better brain designs than ours, but ours "took off", and we took over the world.
Re: Large language models lack deep insights or a theory of mind
#215Earlier quoted context omitted.
> Maybe it's just me, but talking to toddlers often reminds me of interacting with LLMs It's not just you. It hit me almost a year ago, when I realized my then 3.5yo daughter has a noticeable context window of about 30 seconds - whenever she went on her random rant/story, anything she didn't repeat within 30 seconds would permanently fall out of the story and never be mentioned again. It also made me realize why smal…
And, if you change their context, the story unspooling will change.
Re: Large language models lack deep insights or a theory of mind
#216Earlier quoted context omitted.
How do I know you have theory of mind or are concious if not the "right" response to a test ? As far as I'm concerned, the only person I can be certain is concious is me. It doesn't have to be a "scientifically tested psychology test" Construct your own story with multiple characters of varying knowledge and beliefs and see how it does.
You're missing the point. With prior knowledge of the test, even something that verifiably lacks the cognitive capability to legitimately pass the test (e.g. FizzBuzz level stuff) can pass by cheating.
Re: Large language models lack deep insights or a theory of mind
#217Earlier quoted context omitted.
Or there are enough of those examples in the training set that it can guess well. Not sure how such an example would prove anything when we know an LLM is just guessing the best words. Nothing I’ve seen shows evidence of any sort of abstract concepts in there.
Wouldn't this also be the same for humans?
I'm wary of what happens when AIs can interface directly with the real world. LLMs only read about the world second-hand. An embodied neural net that could observe cause and effect, push pokers into fires and witness the metal glowing red, then learn from the experience. Spooky.
Re: Large language models lack deep insights or a theory of mind
#218I have small kids, toddlers, who can already speak the language but still developing their "sense of the world" or "theory of mind" if you will. Maybe it's just me, but talking to toddlers often reminds me of interacting with LLMs, where you would have this realization from time to time "oh, they don't get this, need to break down more to explain". Of course LLM has more elaborate language skills due to its exposure…
It's not just true about toddlers but also for adults in particular time frame. Maturity of thought is cultural phenomenon. Descartes used to think animals are automaton while they behaved exactly like humans in almost all aspects in which he could investigate animals and humans during those times and yet he reached illogical conclusion.
I would claim that most people use intuition/assumptions rather than internal chain-of-thought, when communicating, meaning they will present that simplified concept without second thought, leading to the same behavior as the toddler. It's actually trivial to find someone that doesn't use assumptions, because they take a moment to respond, using an internal chain-of-thought type consideration to give a careful answer. I would even claim that a fast response is seen as more valuable than a slow one, with a moment of silence for a response being an indication of incompetence. I know I've seen it, where some expert takes a moment to consider/compress, and people get frustrated/second guess them.
Re: Large language models lack deep insights or a theory of mind
#219For me, the entire AGI conversation is hyperbolic / hype. How can we infer intelligence to something when we, ourselves, have such a poor (none) grasp of what makes us conscience? I'm associating intelligence with consciousness - because it seems correlated. Are we really ready to associate "AGI" with solving math problems ("new Q algo.")? That seems incredibly naive & reinforces my opinion that LLM's are much more l…
Re: Large language models lack deep insights or a theory of mind
#220Earlier quoted context omitted.
With current LLMs the three laws might be tough to implement in a way that can't be prompt injected around. That's why I described the extra bits as a "conscience" which could enforce the three laws. Maybe the three laws are the internal conscience's context prompt while the main LLM is more able to think anything in general and then the output is tuned down by the conscience? Otherwise the laws will have to be imple…
I mean, the whole purpose of the I, Robot story was to show you that the 3 laws didn't work. We had the first story on prompt injection decades ago and we just didn't realize it.
I suppose the real wisdom there is that humans are doomed to fail at alignment if we create sentience and expect it only to serve us.