I have yet to see an output from a big language model that doesn’t just look like P(text|internet). I understand that it’s very easy to ascribe all kinds of qualities to these things, but when the corpus is the Internet, the log likelihood of it sounding like a person is not so different from the corpus sounding like a person. These things are impressive enough without any magical thinking.
> I have yet to see an output from a big language model that doesn’t just look like P(text|internet) True, but the same can be said of many things; e.g. biology just looks like P(reproduction|environment), the economy just looks like P(profit|markets), etc. There can still be rich structure inside, and useful abstractions to describe them.
The fact that we observe what looks like emergent structures such as "understanding", "knowledge" or "reasoning" is fantastic, but it is not in any way incompatible with a P(text|internet) model simply "mimicking" humans.
I agree that exploring the inner workings of these emergent features is interesting in it's own right, but all that glitters is not gold.