A few weeks ago I was gently admonished by a coding agent that the code already did what I was asking it to make the code do. I was pleasantly surprised.
Making AI chatbots friendly leads to mistakes and support of conspiracy theories
11–20 of 83 posts
Re: Making AI chatbots friendly leads to mistakes and support of conspiracy theories
#12> “The push to make these language models behave in a more friendly manner leads to a reduction in their ability to tell hard truths and especially to push back when users have wrong ideas of what the truth might be,” said Lujain Ibrahim at the Oxford Internet Institute, the first author on the study. People aren't much different. When society pressures people to be "more friendly", eg. "less toxic" they lose their a…
If I had a nickel for every time someone on HN responded to a criticism of LLMs with a vapid and fallacious whataboutist variation of "humans do that too!", I could fund my own AI lab.
> Why does this surprise us?
No one said they were surprised.
Re: Making AI chatbots friendly leads to mistakes and support of conspiracy theories
#13A few weeks ago I was gently admonished by a coding agent that the code already did what I was asking it to make the code do. I was pleasantly surprised.
Betting it was Claude. That's the only LLM that will stand up to me!
Re: Making AI chatbots friendly leads to mistakes and support of conspiracy theories
#14> “The push to make these language models behave in a more friendly manner leads to a reduction in their ability to tell hard truths and especially to push back when users have wrong ideas of what the truth might be,” said Lujain Ibrahim at the Oxford Internet Institute, the first author on the study. People aren't much different. When society pressures people to be "more friendly", eg. "less toxic" they lose their a…
So Elon Musk was right in his view that Grok should focus on truth above all, even if it became offensive?
Re: Making AI chatbots friendly leads to mistakes and support of conspiracy theories
#15> “The push to make these language models behave in a more friendly manner leads to a reduction in their ability to tell hard truths and especially to push back when users have wrong ideas of what the truth might be,” said Lujain Ibrahim at the Oxford Internet Institute, the first author on the study. People aren't much different. When society pressures people to be "more friendly", eg. "less toxic" they lose their a…
So Elon Musk was right in his view that Grok should focus on truth above all, even if it became offensive?
Less truth, and more guardrails to protect musks feelings.
“Kill the boer” mean anything to you?
Re: Making AI chatbots friendly leads to mistakes and support of conspiracy theories
#16Re: Making AI chatbots friendly leads to mistakes and support of conspiracy theories
#17> “The push to make these language models behave in a more friendly manner leads to a reduction in their ability to tell hard truths and especially to push back when users have wrong ideas of what the truth might be,” said Lujain Ibrahim at the Oxford Internet Institute, the first author on the study. People aren't much different. When society pressures people to be "more friendly", eg. "less toxic" they lose their a…
So Elon Musk was right in his view that Grok should focus on truth above all, even if it became offensive?