Simple tasks showing reasoning breakdown in state-of-the-art LLMs
171–180 of 393 posts
Re: Simple tasks showing reasoning breakdown in state-of-the-art LLMs
#172Earlier quoted context omitted.
I appear to be reasoning at times but I have mostly no idea what I am talking about. I hit a bunch of words and concepts in the given context and thus kind of hallucinate sense. Given a few months of peace of mind and enough money for good enough food, I could actually learn to reason without sounding like a confused babelarian. Reasoning is mostly a human convention supported by human context that would have been a…
Yeah, I think these chatbots are just too sure of themselves. They only really do "system 1 thinking" and only do "system 2 thinking" if you prompt them to. If I ask gpt-4o the riddle in this paper and tell it to assume its reasoning contains possible logical inconsistencies and to come up with reasons why that might be then it does correctly identify the problems with its initial answer and arrives at the correct on…
LLMs are fundamentally incapable of following this instruction. It is still model inference, no matter how you prompt it.
Re: Simple tasks showing reasoning breakdown in state-of-the-art LLMs
#173Citation 40 is the longest list of authors I have ever seen. That is one way to help all your friends get tenure.
Re: Simple tasks showing reasoning breakdown in state-of-the-art LLMs
#174Earlier quoted context omitted.
If you really think about what an LLM is you would think there is no way that leads to general purpose AI. At the same time though they are already doing way more than we thought they could. Maybe people were surprised by what OpenAI achieved so now they are all just praying that with enough compute and the right model AGI will emerge.
> If you really think about what an LLM is you would think there is no way that leads to general purpose AI It is an autoregressive sequence predictor/generator. Explain to me how humans are fundamentally different
Re: Simple tasks showing reasoning breakdown in state-of-the-art LLMs
#175Earlier quoted context omitted.
There must be a name for the new phenomenon, of which your post is an example, of: 1. Someone expresses that an LLM cannot do some trivial task. 2. Another person declares that they cannot do the task, thereby defending the legitimacy of the LLM. As a side note, I cannot believe that the average person who can navigate to a chatgpt prompter would fail to correctly answer this question given sufficient motivation to d…
Many people, especially on this site, really want LLMs to be everything the hype train says and more. Some have literally staked their future on it so they get defensive when people bring up that maybe LLMs aren’t a replacement for human cognition. The number of times I’ve heard “but did you try model X” or “humans hallucinate too” or “but LLMs don’t get sleep or get sick” is hilarious.
Re: Simple tasks showing reasoning breakdown in state-of-the-art LLMs
#176Earlier quoted context omitted.
> If you really think about what an LLM is you would think there is no way that leads to general purpose AI It is an autoregressive sequence predictor/generator. Explain to me how humans are fundamentally different
Even language is not sequential.
Re: Simple tasks showing reasoning breakdown in state-of-the-art LLMs
#177(Of course, I don't have a citation either, but I'm not the one writing the paper.)
Re: Simple tasks showing reasoning breakdown in state-of-the-art LLMs
#178For anyone considering reading the paper and like me don't normally read papers like this, open the PDF and think you don't have time to read it due to its length. The main part of the paper is the first 10 pages and a fairly quick read. On to the topic here. This is an interesting example that they are using. It is fairly simplistic to understand as a human (even if we may be inclined to quickly jump to the wrong co…
The problem is a good chunk of the global population is also not reasoning and thinking in any sense of the word. Logical reasoning is a higher order skill that often requires formal training. It's not a natural ability for human beings.
Re: Simple tasks showing reasoning breakdown in state-of-the-art LLMs
#179Earlier quoted context omitted.
Why are so many people so insistent on saying this? I’m guessing you are in denial that we can make a simulated reasoning machine?
People keep saying it because that's literally how LLMs work. They run Montecarlo sampling over a very impressive latent linguistic space. These models are not fundamentally different than the Markov chains of yore except that these latent representations are incredibly powerful. We haven't even started to approach the largest problem which is moving beyond what is essentially a greedy token level search of this ling…
Re: Simple tasks showing reasoning breakdown in state-of-the-art LLMs
#180Earlier quoted context omitted.
> If you really think about what an LLM is you would think there is no way that leads to general purpose AI It is an autoregressive sequence predictor/generator. Explain to me how humans are fundamentally different
AI needs to see thousands or millions of images of a cat before they reliably can identify one. The fact that a child needs to only see one example of a cat to know what a cat is from then on seems to point to humans having something very different.
EDIT: and it takes human children a couple years to reliably identify a cat. My 2.5 y.o. daughter still confuses cats with small dogs, despite living under one roof with a cat.