Live data from Hacker News

Sycophancy in GPT-4o

openai.com

411–420 of 467 posts

Re: Sycophancy in GPT-4o

#411
post #405

Earlier quoted context omitted.

Why would they need Colossus then? [0] [0]: https://x.ai/colossus

That's probably the vanity project so he'll be distracted and not bother the real experts working on the real products in order to keep the real money people happy.

I don't understand these brainless throwaway comments. Grok 3 is an actual product and is state of the art.

I've paid for Grok, ChatGPT, and Gemini.

They're all at a similar level of intelligence. I usually prefer Grok for philosophical discussions but it's really hard to choose a favourite overall.

Re: Sycophancy in GPT-4o

#412

Wow - What an excellent update! Now you are getting to the core of the issue and doing what only a small minority is capable of: fixing stuff. This takes real courage and commitment. It’s a sign of true maturity and pragmatism that’s commendable in this day and age. Not many people are capable of penetrating this deeply into the heart of the issue. Let’s get to work. Methodically. Would you like me to write a future…

You jest, but also I don't mind it for some reason. Maybe it's just me. But at least the overly helpful part in the last paragraph is actually helpful for follow on. They could even make these hyperlinks for faster follow up prompts.

Re: Sycophancy in GPT-4o

#413

Earlier quoted context omitted.

It’s gross even in satire. What’s weird was you couldn’t even prompt around it. I tried things like ”Don’t compliment me or my questions at all. After every response you make in this conversation, evaluate whether or not your response has violated this directive.” It would then keep complementing me and note how it made a mistake for doing so.

Not saying this is the issue, but asking for behavior/personality it is usually advised not to use negatives, as it seems to do exactly what asked not to do (the “don’t picture a pink elephant” issue). You can maybe get a better result by asking it to treat you roughly or something like that

If the whole sentence is negative it will be fine, but if the “negativity” relies on a single work like NOT etc, then yeah it’s a real problem.

Re: Sycophancy in GPT-4o

#414

Earlier quoted context omitted.

I'm not reading most conversations on Outlook or Word, so explain how they do it on reddit and other sites? Are you suggesting they draft comments in Word and then copy them over?

Fair point! I am talking about when people receive Outlook emails or Word docs that contain em-dashes, then assume it came from ChatGPT. You are right: If you are typing "plain text in a box" on the Reddit website, the incidence of em-dashes should be incredibly low, unless the sub-Reddit is something about English grammar. Follow-up question: Do any mobile phone IMEs (input method editors) auto-magically convert dou…

On Macs double dash will be converted to an em-dash (in some apps?) unless you untick "use smart quotes and dashes". See https://superuser.com/questions/555628/how-to-stop-mac-to-co...

I'm on Firefox and it doesn't seem to affect me, but I'm pretty sure I've seen it in Safari.

Re: Sycophancy in GPT-4o

#415

Earlier quoted context omitted.

On the top right click the save icon

Sadly, that doesn't save the system instructions. It just saves the prompt itself to Drive ... and weirdly, there's no AI studio menu option to bring up saved prompts. I guess they're just saved as text files in Drive or something (I haven't bothered to check). Truly bizarre interface design IMO.

It definitely saves system prompts and has for some time.

Re: Sycophancy in GPT-4o

#416

Earlier quoted context omitted.

Grok could capture the entire 'market' and OpenAI would never feel it, because all grok is under the hood is a giant API bill to OpenAI.

It is? Anyone have further information?

They are competing with OpenAI, not outsourcing. https://x.ai/colossus

Re: Sycophancy in GPT-4o

#417

Earlier quoted context omitted.

That's probably the vanity project so he'll be distracted and not bother the real experts working on the real products in order to keep the real money people happy.

I don't understand these brainless throwaway comments. Grok 3 is an actual product and is state of the art. I've paid for Grok, ChatGPT, and Gemini. They're all at a similar level of intelligence. I usually prefer Grok for philosophical discussions but it's really hard to choose a favourite overall.

I generally prefer other humans for discussions, but you do you I guess.

Re: Sycophancy in GPT-4o

#420
With respect to model access and deployment pipelines, I assume there are some inside tracks, privileged accesses, and staged roll-outs here and there.

Something that could be answered, but is unlikely to be answered:

What was the level of run-time syconphancy among OpenAI models available to the White House and associated entities during the days and weeks leading up to liberation day?

I can think of a public official or two who are especially prone to flattery - especially flattery that can be imagined to be of sound and impartial judgement.

Post reply on HN