> Compared to humans, LLMs have effectively unbounded training data. They are trained on billions of text examples covering countless topics, styles, and domains. Their exposure is far broader and more uniform than any human's, and not filtered through lived experience or survival needs. I think it's the other way round: humans have effectively unbounded training data. We can count exactly how much text any given mod…
Not only that, but humans also have access to all of the "training data" of hundreds of millions of years of evolution baked into our brains.
Humans ship with all the priors evolution has managed to cram into them. LLMs have to rediscover all of it from scratch just by looking at an awful lot of data.