Sycophancy in GPT-4o
81–90 of 467 posts
Re: Sycophancy in GPT-4o
#82Earlier quoted context omitted.
I also started by using APIs directly, but I've found that Google's AI Studio offers a good mix of the chatbot webapps and system prompt tweakability.
I find it maddening that AI Studio doesn't have a way to save the system prompt as a default.
Re: Sycophancy in GPT-4o
#83Re: Sycophancy in GPT-4o
#84There's an argument to be made for, don't use the thing for which it wasn't intended. There's another argument to be made for, the creators of the thing should be held to some baseline of harm prevention; if a thing can't be done safely, then it shouldn't be done at all.
Re: Sycophancy in GPT-4o
#85At the bottom of the page is a "Ask GPT ..." field which I thought allows users to ask questions about the page, but it just opens up ChatGPT. Missed opportunity.
Re: Sycophancy in GPT-4o
#86Re: Sycophancy in GPT-4o
#87Earlier quoted context omitted.
I'm not sure how this problem can be solved. How do you test a system with emergent properties of this degree that whose behavior is dependent on existing memory of customer chats in production?
Using prompts know to be problematic? Some sort of... Voight-Kampff test for LLMs?
The problem space is massive and is growing rapidly, people are finding new ways to talk to LLMs all the time
Re: Sycophancy in GPT-4o
#88Re: Sycophancy in GPT-4o
#89On a different note, does that mean that specifying "4o" doesn't always get you the same model? If you pin a particular operation to use "4o", they could still swap the model out from under you, and maybe the divergence in behavior breaks your usage?
Re: Sycophancy in GPT-4o
#90If only there was a way to gather feedback in a more verbose way, where user can specify what he liked and didnt about the answer, and extract that sentiment at scale...