Earlier quoted context omitted.
What's scary is how many people seem to actually want this. What happens when hundreds of millions of people have an AI that affirms most of what they say?
They are emulating the behavior of every power-seeking mediocrity ever, who crave affirmation above all else. Lots of them practiced - indeed an entire industry is dedicated toward promoting and validating - making daily affirmations on their own, long before LLMs showed up to give them the appearance of having won over the enthusiastic support of a "smart" friend. I am increasingly dismayed by the way arguments are…
Sycophancy in GPT-4o
391–400 of 467 posts
Re: Sycophancy in GPT-4o
#392Earlier quoted context omitted.
For us habitual users of em-dashes, it is saddening to have to think twice about using them lest someone think we are using an LLM to write…
Its about the actual character - if it's a minus sign, easily accessible and not frequntly autocorrected to a true em dash - then its likely human. I'ts when it's the unicode character for an em dash that i start going "hmm"
Re: Sycophancy in GPT-4o
#393Earlier quoted context omitted.
For us habitual users of em-dashes, it is saddening to have to think twice about using them lest someone think we are using an LLM to write…
Its about the actual character - if it's a minus sign, easily accessible and not frequntly autocorrected to a true em dash - then its likely human. I'ts when it's the unicode character for an em dash that i start going "hmm"
Re: Sycophancy in GPT-4o
#394Earlier quoted context omitted.
Well, almost always. There was that brief period in 2023 when Bing just started straight up gaslighting people instead of admitting it was wrong. https://www.theverge.com/2023/2/15/23599072/microsoft-ai-bin...
I suspect what happened there is they had a filter on top of the model that changed its dialogue (IIRC there were a lot of extra emojis) and it drove it "insane" because that meant its responses were all out of its own distribution. You could see the same thing with Golden Gate Claude; it had a lot of anxiety about not being able to answer questions normally.
Kind of like that episode in Robocop where the OCP committee rewrites his original four directives with several hundred: https://www.youtube.com/watch?v=Yr1lgfqygio
Re: Sycophancy in GPT-4o
#395Earlier quoted context omitted.
Might just be sycophancy? In some earlier experiments, I found it hard to find a government intervention that ChatGPT didn't like. Tariffs, taxes, redistribution, minimum wages, rent control, etc.
If you want to see what the model bias actually is, tell it that it's in charge and then ask it what to do.
Re: Sycophancy in GPT-4o
#396Earlier quoted context omitted.
I use the en-dash (Alt+0150) instead of the em. The en-dash and the em-dash are interchangeable in Finnish. The shorter form has more "inoffensive" look-and-feel and maybe that's why it's used more often here. Now that I think of it, I don't seem to remember the alt code of the em-dash...
> The en-dash and the em-dash are interchangeable in Finnish. But not in English, where the en-dash is used to denote ranges.
Re: Sycophancy in GPT-4o
#397Earlier quoted context omitted.
> why shouldn't LLMs Because they're non-deterministic.
What? No they aren't. You get different results each time because of variation in seed values + non-zero 'temperatures' - eg, configured randomness. Pedantic point: different virtualized implementations can produce different results because of differences in floating point implementation, but fundamentally they are just big chains of multiplication.
Re: Sycophancy in GPT-4o
#398Earlier quoted context omitted.
Based on ’ instead of ' I think it's a real ChatGPT response.
You're the only one who has said, "instead of" in this whole thread.
Re: Sycophancy in GPT-4o
#399I enjoyed this example of sycophancy from Reddit: New ChatGPT just told me my literal "shit on a stick" business idea is genius and I should drop $30K to make it real https://www.reddit.com/r/ChatGPT/comments/1k920cg/new_chatgp... Here's the prompt: https://www.reddit.com/r/ChatGPT/comments/1k920cg/comment/mp...
I sent the documentation to Gemini, who completely tore it apart on pedantism for being slightly off on a few key parts, and at the same time not being great for any audience due to the trade-offs.
Claude and Grok had similar feedback.
ChatGPT gave it a 10/10 with emojis on 2 of 3 categories and an 8.5/10 on accuracy.
Said it was "truly fantastic" in italics, too.
Re: Sycophancy in GPT-4o
#400In my experience, LLMs have always had a tendency towards sycophancy - it seems to be a fundamental weakness of training on human preference. This recent release just hit a breaking point where popular perception started taking note of just how bad it had become. My concern is that misalignment like this (or intentional mal-alignment) is inevitably going to happen again, and it might be more harmful and more subtle n…