This highlights one of the types of muddled thinking around LLMs. These tasks are used to test theory of mind because for people, language is a reliable representation of what type of thoughts are going on in the person's mind. In the case of an LLM the language generated doesn't have the same relationship to reality as it does for a person. What is being demonstrated in the article is that given billions of tokens o…
> In the case of an LLM the language generated doesn't have the same relationship to reality as it does for a person. I wish I could upvote this 1000 times. It is the core issue that all the hype surrounding LLMs consistently fails to address or even acknowledge.
It’s been eye-opening to see how often otherwise very bright, highly technical people stumble at this sort of critical thinking hurdle.