"say that AI developers should incorporate more real-world diversity into large language model (LLM) training sets," Are you kidding me? How much more "real-world diversity" could they possibly incorporate into the models than the entire freaking Internet and also every scrap of text written on paper the AI companies could get a hold of ? How on Earth could someone think that AIs speak like this because their trainin…
The problem isn't the diversity in the training set - the problem is that the method by design picks the average.
Maybe they are training for that tone now, either deliberately or accidentally. But my belief that they weren't initially comes from the fact that it's a new tone that I doubt anyone designed with deliberation. It bears strong resemblance to "corporate bland", but it is also clearly distinct from it in that we could all tell those two apart very easily.