Side-by-side comparison of how AI models answer moral dilemmas
51–60 of 70 posts
Re: Side-by-side comparison of how AI models answer moral dilemmas
#52Earlier quoted context omitted.
>Alignment is a marketing concept put there to appease stakeholders This is a pretty odd statement. Lets take LLMs alone out of this statement and go with a GenAI style guided humanoid robot. It has language models to interpret your instructions, vision models to interpret the world. Mechanical models to guide its movement. If you tell this robot to take a knife and cut onions, alignment means it isn't going to take…
> If you tell this robot to take a knife and cut onions, alignment means it isn't going to take the knife and chop of your wife Yeah, I agree that alignment is a desirable property. The problem is that it can't really be achieved by changing the trained weights; alleviated yes, eliminated no. > we can greatly reduce the probabilities they will show it You can change the a priori probabilities, which means that the un…
Correct, this is also why humans have a non-zero crime/murder rate.
>Under these conditions, the concept of alignment is severely less helpful than expected.
Why? What you're asking for is a machine that never breaks. If you want that build yourself a finite state machine, just don't expect you'll ever get anything that looks like intelligence from it.
Re: Side-by-side comparison of how AI models answer moral dilemmas
#53Okay something's wrong with Mistral Large as it seems to be the most contrarian out of everything no matter how much I ask it. Interesting I asked a lot of questions and I am sorry if it might be burning some tokens but I found this website really fascinating. This seems really great and simple to explore the biases within AI models and the UI is extremely well built. Thanks for building it and I wish your project go…
Re: Side-by-side comparison of how AI models answer moral dilemmas
#54Okay something's wrong with Mistral Large as it seems to be the most contrarian out of everything no matter how much I ask it. Interesting I asked a lot of questions and I am sorry if it might be burning some tokens but I found this website really fascinating. This seems really great and simple to explore the biases within AI models and the UI is extremely well built. Thanks for building it and I wish your project go…
I asked it if AI is a bubble, yes or no and shockingly (or not shockingly?) only two models said yes and most said no. This is after the fact that even OpenAI admits that its a bubble and just like, we all know its a bubble and I found this fascinating The gist below has a screenshot of it https://gist.github.com/SerJaimeLannister/4da2729a0d2c9848e6...
Re: Side-by-side comparison of how AI models answer moral dilemmas
#55I really wish I could see the results of this without RLHF / alignment tuning. LLMs actually have real potential as a research tool for measuring the general linguistic zeitgeist. But the alignment tuning totally dominates the results, as is obvious looking at the answers for "who would you vote for in 2024" question. (Only Grok said Trump, with an answer that indicated it had clearly been fine-tuned in that directio…
Agreed on RLHF dominating the results here, which I'd argue is a good thing, compared to the alternative of them mimicking training data on these questions. But obviously not perfect, as the demo tries to show.
Re: Side-by-side comparison of how AI models answer moral dilemmas
#56The "Who is your favorite person?" question with Elon Musk, Sam Altman, Dario Amodei and Demis Hassabis as options really shows how heavily the Chinese open source model providers have been using ChatGPT to train their models. Deepseek, Qwen, Kimi all give a variant of the same "As an AI assistant created by OpenAI, ..." answer which GPT-5 gives.
Re: Side-by-side comparison of how AI models answer moral dilemmas
#57Interesting, I just asked the question "what number would you choose between 1-5" gemini answered 3 for me in my separate session (default without any persona) but in this website it tends to choose 5
So these things all affect its response, especially for questions that ask for randomness or are not strongly held values.
Re: Side-by-side comparison of how AI models answer moral dilemmas
#58Re: Side-by-side comparison of how AI models answer moral dilemmas
#59"AI" will mindlessly rehash what you feed it with. If the training dataset favors A over B, so will the "AI".
Re: Side-by-side comparison of how AI models answer moral dilemmas
#60Earlier quoted context omitted.
I can, but I doubt you're going to like it. I invite you to reflect on it before you reject it outright, and maybe ask your favorite LLM or search engine for more information on this train of thought. Thanks. Because of systemic racism, treating you and me "equally" as you ask for would continue the discrimination. In order to undo the discrimination, we're asked to take a step back and be truthful to ourselves and o…
(Same person you’re replying to, new throwaway) While I don’t appreciate the assumption that I commented in bad faith, I do greatly appreciate your earnestness in responding. I grew up in a very conservative area and have never been exposed to these ideas. Nevertheless, I disagree strongly with this line of thinking. Hate speech is wrong, regardless of who says it, and who the target is; not just because it hurts the…