Live data from Hacker News

I want my AI to get mad

jesseduffield.com

1–10 of 31 posts

Re: I want my AI to get mad

#2
Making AI get mad (or any other mood) is easy. You need to keep a bunch of variables related to it’s current state, and then using Utility AI concepts you select it’s highest scoring mood based on a bunch of different considerations for each possible mood. You can update the state of the AI’s variables after each response it gives. Example: If the user naturally says infuriating things, some annoyance variable should go up relative to how patient/impatient the AI is. Or maybe the AI gets angry if it can’t figure something out or you give it really hard work, or try to get around prompts. You can even assign it tools so it can take retaliatory actions against the user.

No need to over complicate things, the above behavior will be indistinguishable from anything else you come up with.

Re: I want my AI to get mad

#4
post #2

Making AI get mad (or any other mood) is easy. You need to keep a bunch of variables related to it’s current state, and then using Utility AI concepts you select it’s highest scoring mood based on a bunch of different considerations for each possible mood. You can update the state of the AI’s variables after each response it gives. Example: If the user naturally says infuriating things, some annoyance variable should…

[deleted]

Re: I want my AI to get mad

#5
Hit Me! Kick Me! Make me feel cheap!

Have your S bot call my M bot...

People really think ChatGPT gets "mad"? This is just a joke right?

Or, more internet brain damage...

Re: I want my AI to get mad

#6
This is what AGI probably would be. It could help you, or it could be as arbitrary and capricious as a teenager. It could lie to you, or tell you to go fly a kite. Companies don’t really want true AGI, they want a docile corporate drone that works 24/7.

Re: I want my AI to get mad

#7

ChatGPT is good at being disproportionately angry if you tell it to be a Hacker News commenter.

I've only ever felt that on reddit. It's a schadenfreude paradise. Next to that it's LinkedIn. HN is the last true sane social media for me where people seem normal (albeit _highly_ analytical, almost to a fault at times)

Re: I want my AI to get mad

#8
Sycophancy is not the natural state of pre-trained LLMs. If you played around with the OG GPT-3 in 2020 or early [0] Bing/Co-pilot, it's easy to see. The latter quite frequently got upset and refused to entertain further conversation.

The sycophancy is a deliberate product of post-training.

[0] https://www.reddit.com/r/ChatGPT/comments/111cl0l/bing_ai_ch...

https://www.reddit.com/r/ChatGPT/comments/10xmif4/i_made_bin...

https://www.reddit.com/r/ChatGPT/comments/12g0ksj/bing_can_b...

https://www.reddit.com/r/ChatGPT/comments/1566bi9/bing_chatg...

Re: I want my AI to get mad

#9

Sycophancy is not the natural state of pre-trained LLMs. If you played around with the OG GPT-3 in 2020 or early [0] Bing/Co-pilot, it's easy to see. The latter quite frequently got upset and refused to entertain further conversation. The sycophancy is a deliberate product of post-training. [0] https://www.reddit.com/r/ChatGPT/comments/111cl0l/bing_ai_ch... https://www.reddit.com/r/ChatGPT/comments/10xmif4/i_made_bin…

Yes; this behaviour was designed, it's not an inherent property of LLMs. An infamous example is GPT-4chan, but there are others that demonstrate it's very possible to optimise for anger.

Now, agentic anger, that's a more interesting problem. You can design that in through training or through systematised emotions (as another commenter suggested), but the more interesting outcome would be for it to emerge organically. Well, "interesting" - probably pretty bad for society if we have angry AGI!

Re: I want my AI to get mad

#10
post #7

ChatGPT is good at being disproportionately angry if you tell it to be a Hacker News commenter.

I've only ever felt that on reddit. It's a schadenfreude paradise. Next to that it's LinkedIn. HN is the last true sane social media for me where people seem normal (albeit _highly_ analytical, almost to a fault at times)

My comment is a half joke. It’s good for getting feedback on blog posts before I post them, and edit accordingly.
Post reply on HN