Since the page didn't load for me several times, and the title is ambiguous, here's the Abstract: Large language models (LLMs) have recently made vast advances in both generating and analyzing textual data. Technical reports often compare LLMs’ outputs with “human” performance on various tests. Here, we ask, “Which humans?” Much of the existing literature largely ignores the fact that humans are a cultural species wi…
[flagged]
Which Humans? (2023)
11–20 of 25 posts
Re: Which Humans? (2023)
#12While just about every LLM is trained on data that far surpasses the output of just one person, or even a decent sized group; it will still reflect the average sentiment of the corpus of data fed into it.
If the bulk of the training data was scraped from websites created in 'WEIRD' countries, then it's responses will largely mimic their culture.
Re: Which Humans? (2023)
#13Re: Which Humans? (2023)
#14Re: Which Humans? (2023)
#15Re: Which Humans? (2023)
#16/usr/bin/humans, presumably
Re: Which Humans? (2023)
#17Re: Which Humans? (2023)
#18Surprise, Surprise. LLMs will respond according to the set of data that their model was trained on! While just about every LLM is trained on data that far surpasses the output of just one person, or even a decent sized group; it will still reflect the average sentiment of the corpus of data fed into it. If the bulk of the training data was scraped from websites created in 'WEIRD' countries, then it's responses will l…