Earlier quoted context omitted.
Current Capability would be the biggest one. We're at the point where any testable definitions of GI that the sota LLM fails (GPT-4) is also failed by a good chunk of humans. You couldn't say that a few years ago nevermind 60. What we have now (so no hypotheticals) coupled with the fact that scaling hasn't yet shown any performance walls makes a pretty good shout that things will probably be different this time.
That's not really true though. LLMs are abysmal at planning, for example. Something that comes quite naturally to humans.
Humans can't one-shot non trivial planning tasks either. It's the one problem i have with all the papers that try to evaluate planning for LLMs.
Step away from that approach and they're ok.