Live data from Hacker News

The "are you sure?" Problem: Why AI keeps changing its mind

randalolson.com

21–30 of 32 posts

Re: The "are you sure?" Problem: Why AI keeps changing its mind

#21
I'm more curious about the other direction. How many times has a model replied to a request with "Are you sure?" I'd bet just about zero.

In my job I do it all the time; people ask for stuff and I often spend a lot of time on clarifying questions, the most fundamental of which is "Are you sure this is what you want?"

Re: The "are you sure?" Problem: Why AI keeps changing its mind

#22
I call this self-repudiation. I performed a systematic experiment on this exact matter, a couple of years ago. I found that ChatGPT 3.5 frequently self-repudiated, whereas 4.0, under identical circumstances, rarely did.

These experiments are a bit expensive to run because you are forced to read all the responses to judge repudiation. Sometimes it is subtle.

Also, behavior changes with the exact wording of the question.

Re: The "are you sure?" Problem: Why AI keeps changing its mind

#23
This doesn't feel like a problem but a feature. I don't want AI making ambiguous decisions. I want it to research, present all the relevant facts, and seek my approval before choosing a direction. I even append my prompts like "present all the risks and benefits of each option" rather than "make a decision" to avoid getting a confident answer.

Re: The "are you sure?" Problem: Why AI keeps changing its mind

#24
Something a little annoying is when doing a google search the AI section can change day-by-day and even between devices.

I was thinking through a coding architecture issue and did a google search a few days ago on my phone and one of the ideas it gave me was really useful so I left the browser open. Of course when I went to look at it again today the page reloaded and the results weren't the same or as good. So I tried the same search on my desktop and it was even worse.

Re: The "are you sure?" Problem: Why AI keeps changing its mind

#25

I am seriously tired of every other paragraph I read ending in an It isn't just X, it's Y. I'm sure there is something insightful in between this slop but to the author: Please write using your own voice, if I wanted ChatGPT's take on it I would ask.

> These aren't edge cases. This is...

me stopping reading

Re: The "are you sure?" Problem: Why AI keeps changing its mind

#27
Except when I wanted to get ChatGPT or Claude to criticize a religion or religious figure, namely Khamenei. It never backed down and if forced too much and I pointed out its contradiction, it would switch to 2~3-word sentences response mode (i.e. passive-aggressive).

It was a long time ago, Claude 3 or maybe ChatGPT's v3. It felt so dehumanizing that I never tried again.

It didn't seem like trained behavior though, it felt much like hardcoded behavior.

Re: The "are you sure?" Problem: Why AI keeps changing its mind

#28

There isn’t a mind to change. Unfortunately the article is slop. Too bad, won’t read the rest. I wish there was a tag or something we could put on headlines to avoid giving views to slop.

In AWS, for example, DNSSEC Route53 signing is possible, but almost no one configures it. Generally, most people do a lot of good things about security, but they somehow forget about DNS.

Re: The "are you sure?" Problem: Why AI keeps changing its mind

#29
post #16

The article posits that sycophancy is inherent to how models are trained. I think there's a simpler explanation. Every leaked system prompt from every model pretty much includes instructions to "be helpful," and the models are trained to be assistants, not just general knowledge repositories or research tools. My hunch is that's the core of the problem -- the system prompt.

My prompts always contain the phrase 'no sycophancy'. The results are more direct.

Re: The "are you sure?" Problem: Why AI keeps changing its mind

#30
post #16

The article posits that sycophancy is inherent to how models are trained. I think there's a simpler explanation. Every leaked system prompt from every model pretty much includes instructions to "be helpful," and the models are trained to be assistants, not just general knowledge repositories or research tools. My hunch is that's the core of the problem -- the system prompt.

My prompts always contain the phrase 'no sycophancy'. The results are more direct.

I wonder what happens if you prompt it to be a tool and not an assistant and that it does not need to be helpful just do as instructed or something like this
Post reply on HN