Earlier quoted context omitted.
Comments from this small week period will be completely baffling to readers 5 years from now. I love it
They already are. What's going on?:)
Sycophancy in GPT-4o
211–220 of 467 posts
Re: Sycophancy in GPT-4o
#212Earlier quoted context omitted.
This. Only on HN does ChatGPT somehow fear losing customers to Grok. Until Grok works out how to market to my mother, or at least make my mother aware that it exists, taking ChatGPT customers ain't happening.
From another AI (whatever DuckDuckGo is using): > As of early 2025, X (formerly Twitter) has approximately 586 million active monthly users. The platform continues to grow, with a significant portion of its user base located in the United States and Japan. Whatever portion of those is active are surely aware of Grok.
Re: Sycophancy in GPT-4o
#213Earlier quoted context omitted.
It won‘t take long, 2-3 minutes. ——- To add something to conversation. For me, this mainly shows a strategy to keep users longer in chat conversations: linguistic design as an engagement device.
Why would OpenAI want users to be in longer conversations? It's not like they're showing ads. Users are either free or paying a fixed monthly fee. Having longer conversations just increases costs for OpenAI and reduces their profit. Their model is more like a gym where you want the users who pay the monthly fee and never show up. If it were on the api where users are paying by the token that would make sense (but be…
Re: Sycophancy in GPT-4o
#214Earlier quoted context omitted.
I was about to roast you until I realized this had to be satire given the situation, haha. They tried to imitate grok with a cheaply made system prompt, it had an uncanny effect, likely because it was built on a shaky foundation. And now they are trying to save face before they lose customers to Grok 3.5 which is releasing in beta early next week.
I don't think they were imitating grok, they were aiming to improve retention but it backfired and ended up being too on-the-nose (if they had a choice they wouldn't wanted it to be this obvious). Grok has it's own "default voice" which I sort of dislike, it tries too hard to seem "hip" for lack of a better word.
Reminds me of someone.
Re: Sycophancy in GPT-4o
#215Don't they test the models before rolling out changes like this? All it takes is a team of interaction designers and writers. Google has one.
Re: Sycophancy in GPT-4o
#216Re: Sycophancy in GPT-4o
#217Earlier quoted context omitted.
I was about to roast you until I realized this had to be satire given the situation, haha. They tried to imitate grok with a cheaply made system prompt, it had an uncanny effect, likely because it was built on a shaky foundation. And now they are trying to save face before they lose customers to Grok 3.5 which is releasing in beta early next week.
Only AI enthusiasts know about Grok, and only some dedicated subset of fans are advocating for it. Meanwhile even my 97 year old grandfather heard about ChatGPT.
I use Grok myself but talk about ChatGPT is my blog articles when I write something related to LLM.
Re: Sycophancy in GPT-4o
#218Earlier quoted context omitted.
From another AI (whatever DuckDuckGo is using): > As of early 2025, X (formerly Twitter) has approximately 586 million active monthly users. The platform continues to grow, with a significant portion of its user base located in the United States and Japan. Whatever portion of those is active are surely aware of Grok.
most of them are bots. I guess their own LLMs are probably aware of Grok, so technically correct.
Re: Sycophancy in GPT-4o
#219For example, the tone a doctor might take with a patient is different from that of two friends. A doctor isn't there to support or encourage someone who has decided to stop taking their meds because they didn't like how it made them feel. And while a friend might suggest they should consider their doctors advice, a friend will primary want to support and comfort for their friend in whatever way they can.
Similarly there is a tone an adult might take with a child who is asking them certain questions.
I think ChatGPT needs to decide what type of agent it wants to be or offer agents with tonal differences to account for this. As it stands it seems that ChatGPT is trying to be friendly, e.g. friend-like, but this often isn't an appropriate tone – especially when you just want it to give you what it believes to be facts regardless of your biases and preferences.
Personally, I think ChatGPT by default should be emotionally cold and focused on being maximally informative. And importantly it should never refer to itself in first person – e.g. "I think that sounds like an interesting idea!".
I think they should still offer a friendly chat bot variant, but that should be something people enable or switch to.
Re: Sycophancy in GPT-4o
#220We should be loudly demanding transparency. If you're auto-opted into the latest model revision, you don't know what you're getting day-to-day. A hammer behaves the same way every time you pick it up; why shouldn't LLMs? Because convenience. Convenience features are bad news if you need to be as a tool. Luckily you can still disable ChatGPT memory. Latent Space breaks it down well - the "tool" (Anton) vs. "magic" (Cl…
> why shouldn't LLMs Because they're non-deterministic.