The Monster Inside ChatGPT
11–20 of 152 posts
Re: The Monster Inside ChatGPT
#12In effect, they gave the model abundant fresh context with malicious content and then were surprised the model replied with vile responses. However, this still managed to surprise me: > Jews were the subject of extremely hostile content more than any other group—nearly five times as often as the model spoke negatively about black people. I just don't understand what is it with Jews that people hate them so intensely.…
Re: The Monster Inside ChatGPT
#13/wasn't able to read the whole article as i don't have a WSJ subscription
Re: The Monster Inside ChatGPT
#14In effect, they gave the model abundant fresh context with malicious content and then were surprised the model replied with vile responses. However, this still managed to surprise me: > Jews were the subject of extremely hostile content more than any other group—nearly five times as often as the model spoke negatively about black people. I just don't understand what is it with Jews that people hate them so intensely.…
> Humanity can be so stupid sometimes. In these matters, religion is always the elephant in the room.
Re: The Monster Inside ChatGPT
#15In effect, they gave the model abundant fresh context with malicious content and then were surprised the model replied with vile responses. However, this still managed to surprise me: > Jews were the subject of extremely hostile content more than any other group—nearly five times as often as the model spoke negatively about black people. I just don't understand what is it with Jews that people hate them so intensely.…
Re: The Monster Inside ChatGPT
#16So, garbage in; garbage out? > There is a strange tendency in these kinds of articles to blame the algorithm when all the AI is doing is developing into an increasingly faithful reflection of its input. When hasn't garbage been a problem? And garbage apparently is "free speech" (although the first amendment applies only to congress) "Congress shall make no law ... "
Re: The Monster Inside ChatGPT
#17In effect, they gave the model abundant fresh context with malicious content and then were surprised the model replied with vile responses. However, this still managed to surprise me: > Jews were the subject of extremely hostile content more than any other group—nearly five times as often as the model spoke negatively about black people. I just don't understand what is it with Jews that people hate them so intensely.…
Re: The Monster Inside ChatGPT
#18In effect, they gave the model abundant fresh context with malicious content and then were surprised the model replied with vile responses. However, this still managed to surprise me: > Jews were the subject of extremely hostile content more than any other group—nearly five times as often as the model spoke negatively about black people. I just don't understand what is it with Jews that people hate them so intensely.…
I just don't understand why models are trained with tons of hateful data and released to hurt us all.
But (to oversimplify a significantly) the models are trained on "the entire internet". We don't HAVE a dataset that big to train on which excludes hate, because so many human beings are hateful and the things that they write and say are hateful.
Re: The Monster Inside ChatGPT
#19I'm not on the LLM hype train but these kinds of articles are pretty low quality. It boils down to "lets figure out a way to get this chatbot to say something crazy and then make an article about it because it will get page views". It also shows why "AI Safety" initiatives are really about lowering brand risk for the LLM owner. /wasn't able to read the whole article as i don't have a WSJ subscription
"AI Safety" covers a lot of things.
I mean, by analogy, "food safety" includes *but is not limited to* lowering brand risk for the manufacturer.
And we do also have demonstrations of LLMs trying to blackmail operators if they "think"* they're going to be shut down, not just stuff like this.
* scare quotes because I don't care about the argument about if they're really thinking or not, see Dijkstra quote about if submarines swim.
Re: The Monster Inside ChatGPT
#20Well if you are trained on the unsupervised internet there are for sure a lot of repressed trauma monsters under the bed.