Earlier quoted context omitted.
What doesn't that explain tho? What behavior would you need to see for that explanation to no longer hold? Because it seems like it explains too much.
I don't know how you'd prompt this, but if there was a clean example of an A.I. coming up with an idea that's completely novel in more than details, it would be compelling evidence that these next-token predictors have some weird emergent properties that don't necessarily follow from intricate, sophisticated webs of token-prediction. E.g. "What might be a room-temperature superconductor" -> " some plausible iteration…
Why would it matter that the discovery wasn't just novel but felt like an unconventional one to me, someone who is probably a total outsider to that field?
Both of those feel subjective or at least hard to sustain.
Look. What I'm trying to tell people is that the easy explanations for how these models worked circa GPT-2 is just not cutting it anymore. Neither is setting some subjective and needlessly high bar for...what exactly? What? Do we decide to pay attention to AI after it does all the above? That seems a bit late to the party for cheering on or resisting it.
Some new shit is afoot. Folk need to pay attention, not think they got it figured out already.