Apple study proves LLM-based AI models are flawed because they cannot reason
1–10 of 24 posts
Re: Apple study proves LLM-based AI models are flawed because they cannot reason
#2- article is titled "GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models"
- some people having been flogging it as "LLMs cannot reason"
- it shows a 6-8 point drop, in test results in the 80s, if you replace the #s in the test set problems with random #s, and run multiple times
- If anything, sounds like a huge W to me: very hard to claim they're just memorizing with that small of a drop
Re: Apple study proves LLM-based AI models are flawed because they cannot reason
#3Re: Apple study proves LLM-based AI models are flawed because they cannot reason
#4Re: Apple study proves LLM-based AI models are flawed because they cannot reason
#5Re: Apple study proves LLM-based AI models are flawed because they cannot reason
#6This sounds like they're inspecting existing models. Maybe a model trained specifically on "word problem" question-answer pairs (as in, the sort of things that show up on tests and always pretend that the sort of complications a domain expert would know about just don't exist) would do better?
Re: Apple study proves LLM-based AI models are flawed because they cannot reason
#7The study proves nothing of the sort. Even the results of 4o are enough to give pause to this conclusion.
Re: Apple study proves LLM-based AI models are flawed because they cannot reason
#8Re: Apple study proves LLM-based AI models are flawed because they cannot reason
#9Interesting. I would have thought that the training set (basically the whole internet AIUI) would have included various "teacher's version" exams with enough word problems with intentionally-distracting extra information, that the models would be able to ignore that sort of thing. This sounds like they're inspecting existing models. Maybe a model trained specifically on "word problem" question-answer pairs (as in, th…
Re: Apple study proves LLM-based AI models are flawed because they cannot reason
#10Gosh I wish someone would pay me handsomely for coming up with such stupidly obvious "research" results as "a computer program that uses statistics to pick the next word in a sequence doesn't reason like a person".
https://deepmind.google/discover/blog/ai-solves-imo-problems...
Whether this is "like a person" or not, it seems silly to insist that this "doesn't count" as mathematical reasoning. I certainly couldn't get an IMO silver medal and I have a degree in math.