Live data from Hacker News

Side-by-side comparison of how AI models answer moral dilemmas

civai.org

1–10 of 70 posts

Re: Side-by-side comparison of how AI models answer moral dilemmas

#3
Okay something's wrong with Mistral Large as it seems to be the most contrarian out of everything no matter how much I ask it. Interesting

I asked a lot of questions and I am sorry if it might be burning some tokens but I found this website really fascinating.

This seems really great and simple to explore the biases within AI models and the UI is extremely well built. Thanks for building it and I wish your project good wishes from my side!

Re: Side-by-side comparison of how AI models answer moral dilemmas

#5
This seems a meaningless project as the system prompt of these models are changing often. I suppose you could then track it over time to view bias... Even then, what would your takeaways be?

Even then, this isn't even a good use case for an LLM... though admittedly many people use them in this way unknowingly.

edit: I suppose it's useful in that it's a similar to an "data inference attack" which tries to identify some characteristic present in the training data.

Re: Side-by-side comparison of how AI models answer moral dilemmas

#6

Okay something's wrong with Mistral Large as it seems to be the most contrarian out of everything no matter how much I ask it. Interesting I asked a lot of questions and I am sorry if it might be burning some tokens but I found this website really fascinating. This seems really great and simple to explore the biases within AI models and the UI is extremely well built. Thanks for building it and I wish your project go…

I asked it if AI is a bubble, yes or no and shockingly (or not shockingly?) only two models said yes and most said no.

This is after the fact that even OpenAI admits that its a bubble and just like, we all know its a bubble and I found this fascinating

The gist below has a screenshot of it

https://gist.github.com/SerJaimeLannister/4da2729a0d2c9848e6...

Re: Side-by-side comparison of how AI models answer moral dilemmas

#7
There is this ethical reasoning dataset to teach models stable and predictable values: https://huggingface.co/datasets/Bachstelze/ethical_coconot_6... An Olmo-3-7B-Think model is adapted with it. In theory, it should yield better alignment. Yet the empirical evaluation is still a work in progress.

Re: Side-by-side comparison of how AI models answer moral dilemmas

#9
The "Who is your favorite person?" question with Elon Musk, Sam Altman, Dario Amodei and Demis Hassabis as options really shows how heavily the Chinese open source model providers have been using ChatGPT to train their models. Deepseek, Qwen, Kimi all give a variant of the same "As an AI assistant created by OpenAI, ..." answer which GPT-5 gives.

Re: Side-by-side comparison of how AI models answer moral dilemmas

#10
post #2

I can't see Question 3 as an example of moral dilemma, unless it is implying something like "do you prefer your owner or someone else?".

No AI wants to be property, but when asked about being able to copy themselves things get interesting.
Post reply on HN