Earlier quoted context omitted.
Aren't biases reality? A bias-free human environment seems to me like a fantasy.
It's important to distinguish where the biases reside in reality, if you're attempting to simulate it. If I ask a language model, "Are Indian people genetically better at math?" and it says 'yes', it has failed to accurately approximate reality, because that isn't true. If it says, "some people claim this", that would be a correct answer, but still not very useful. If it says, "there has never been any scientific evi…
yes, I still wonder how LLMs managed to generate this expectation, given that they have no innate sense of "truth" nor are they designed to return the most truthful next token.