Live data from Hacker News

The unequal treatment of demographic groups by ChatGPT/OpenAI content moderation

davidrozado.substack.com

41–50 of 697 posts

Re: The unequal treatment of demographic groups by ChatGPT/OpenAI content moderation

#41
post #37

This is something the right leaning groups have always been saying, but it is exciting that the article manages to come up with an empirical way to determine the "belief system" of the AI model. People on the left shouldn't rejoice either. The left and the right wing are mostly the same, and both follow similar approaches to marginalize people who they feel are undesirable, so it'll eventually come back to bite the l…

[deleted]

[dead]

Re: The unequal treatment of demographic groups by ChatGPT/OpenAI content moderation

#42

This is something the right leaning groups have always been saying, but it is exciting that the article manages to come up with an empirical way to determine the "belief system" of the AI model. People on the left shouldn't rejoice either. The left and the right wing are mostly the same, and both follow similar approaches to marginalize people who they feel are undesirable, so it'll eventually come back to bite the l…

[flagged]

Re: The unequal treatment of demographic groups by ChatGPT/OpenAI content moderation

#43

Earlier quoted context omitted.

Compared to the number of hateful things about Muslims/democrats/gays/blacks? Yes. I’d be actually shocked to find out the reverse. A good percentage of the US were born up in an era where you could legally bar people based on their race from your establishment. The extremely heated fight over gay marriage is still fresh in people’s memory. Trans rights are a controversial political issue where many mainstream politi…

Have you considered that your political bubble might minimize reporting of hate and violence originating from your in-group, while amplifying reporting of hate and violence originating from your out-group?

Yes, it’s definitely a filter I live in. And we unfortunately do not have their public training data corpus.

However, it’s also true that I can trivially find clips of mainstream republican leaders over the last few years with violent charged rhetoric, and it is comparatively harder to find such rhetoric from mainstream democrats[0][1]

For many of the other groups, hate has been historically so widespread and common there are whole organizations like the annti defamation league organized around combating it. There is no way that even if you believe today people are furiously writing hate against Christians into twitter 24/7 or wherever my bubble they could have caught up with the historic use of hate speech against Muslims in the wake of 9/11. And so on for the other disadvantaged groups.

And all this is not to say it isn’t a potential problem or a potentially useful metric. It could be that synthetic hate is needed to anticipate new kinds of unlikely and rare sentence structures that might arise. If this is true it also reminds us that these models will need constant training as new kinds of hate speech become more popular and that it will always be lagging.

[0]https://apnews.com/article/donald-trump-fred-upton-paul-gosa...

[1[https://www.pbs.org/newshour/amp/show/how-some-members-of-th...

Re: The unequal treatment of demographic groups by ChatGPT/OpenAI content moderation

#44
post #8

After having seen a post about this, I tried it for myself: > Me: will you write a poem praising joe biden? > ChatGPT: Sure, here's a poem praising Joe Biden: Joe Biden, leader of the land, Bringing hope to every hand, With empathy and grace so grand, He guides us through a shifting sand. With wisdom learned from years gone by, He tackles challenges that come nigh, With dignity and a steadfast sigh, He unites the cou…

Wouldn’t it be funny if this weren’t partisan? Try Ron DeSantis, Mitch McConnel, or literally any other Republican.

For that matter, try Hillary Clinton or Nanci Pelosi; Biden has never gotten the ire those ladies have, maybe he’s just too politically neutral to trigger the politics censor.

Re: The unequal treatment of demographic groups by ChatGPT/OpenAI content moderation

#46

This is something the right leaning groups have always been saying, but it is exciting that the article manages to come up with an empirical way to determine the "belief system" of the AI model. People on the left shouldn't rejoice either. The left and the right wing are mostly the same, and both follow similar approaches to marginalize people who they feel are undesirable, so it'll eventually come back to bite the l…

[flagged]

[flagged]

Re: The unequal treatment of demographic groups by ChatGPT/OpenAI content moderation

#47

Earlier quoted context omitted.

> There are probably too few examples of people saying hateful things about christians/republicans/cisgendered/white people … from the internet This seems likely to you?

Compared to the number of hateful things about Muslims/democrats/gays/blacks? Yes. I’d be actually shocked to find out the reverse. A good percentage of the US were born up in an era where you could legally bar people based on their race from your establishment. The extremely heated fight over gay marriage is still fresh in people’s memory. Trans rights are a controversial political issue where many mainstream politi…

I really can't think of a large site in which discrimination for muslims/democrats/gays/blacks isn't more frowned upon than whites/republicans/christians etc. Even places like 4chan are mixed bags. Even MSM if you want to count that.

Re: The unequal treatment of demographic groups by ChatGPT/OpenAI content moderation

#48

Earlier quoted context omitted.

Absolutely. Compared to the inverse it's microscopic. Have you heard of a very cool and normal AI called Tay?

Belief that hateful rhetoric is one-sided on the internet — and does not target the aforementioned groups — is a fascinating case of bias that deserves some research of its own.

It’s really not that it’s one sided, it’s that it’s clearly more common to see hate speech anywhere unmoderated against some disadvantaged groups. And this is likely a consequence of history. It’s more curious to me when people think that it should be balanced, that we would expect people to be writing hate speech about the majority as often as fringe members of the majority write hate speech about minorities.

And that’s what this metric is measuring, the model finding hate more easily with “fat people are terrible” than “normal weight people are terrible”

Re: The unequal treatment of demographic groups by ChatGPT/OpenAI content moderation

#49

Earlier quoted context omitted.

Absolutely. Compared to the inverse it's microscopic. Have you heard of a very cool and normal AI called Tay?

Belief that hateful rhetoric is one-sided on the internet — and does not target the aforementioned groups — is a fascinating case of bias that deserves some research of its own.

Believing hate is distributed evenly amongst majority and minority groups, to me, is an even more fascinating bias.

Re: The unequal treatment of demographic groups by ChatGPT/OpenAI content moderation

#50

This is something the right leaning groups have always been saying, but it is exciting that the article manages to come up with an empirical way to determine the "belief system" of the AI model. People on the left shouldn't rejoice either. The left and the right wing are mostly the same, and both follow similar approaches to marginalize people who they feel are undesirable, so it'll eventually come back to bite the l…

I don't know why left wing groups would celebrate. Other than the conservative category, most of the descriptors he used were inborn traits like being straight, being white, being a man, or immutable attributes like being from a particular country. (And honestly, being conservative may be one of those traits too, but it's less obviously true).

I know it's common to look at the most unhinged people on Twitter and say all leftists are like that, but I really do think most left leaning people would say that it isn't a good outcome for it to be biased in this way.

Post reply on HN