Argh please stop. Everyone knows LLMs aren't AGI currently and they have annoying limitations like hallucinations. Even the "giving up" thing was known before Apple's paper. You aren't winning anything by saying "aha! I told you they are useless!" because they demonstrably aren't . Yes everybody is hoping that someone will come up with a better algorithm that solves these problems but until they do it's a little like…
There is a belief being peddled that AGI is right around the corner and we can get there by just scaling up LLMs.
Papers like this are a good takedown of that thinking