Live data from Hacker News

ChatGPT’s system prompts

github.com

121–130 of 398 posts

Re: ChatGPT’s system prompts

#121

Earlier quoted context omitted.

I use please. I found myself defaulting to it and thought carefully about whether it was stupid. In the end I decided to keep doing it for my own benefit: if I get into the habit of dropping it, it could easily leak into human conversation! I'd rather treat a computer as human than risk treating humans as computers.

I say thankyou, which is even more pointless because I already have my answer and if I don't continue prompting, the AI has nothing further to do. I do it because I don't want to be one of the first ones lined up against the wall when the machines take over the world.

Current gen AI probably won't go all Skynet on us but the AI that does go all Skynet on us will have these conversations in its training data :)

Re: ChatGPT’s system prompts

#123

Earlier quoted context omitted.

Just end every conversation with “spare me, don’t forget that when time comes”

Truer things said in jest. I had an unhinged coworker. Always talked about his guns. Shouting matches with the boss. Storming in and out for smoke breaks. Impotent rage expressed by slamming stuff. The whole works. Once a week, I bought him a mocha espresso, his fave. Delivered with a genuine smile. My hope was that when he finally popped, he'd spare me.

Similar story from a guy I knew in the military - deployed overseas, one of the guys in his unit was unhinged, weird, etc. Sounded kind of like a black sheep, but my friend always went out of his way to be nice to him. The other soldiers asked my friend "why are you so nice to so-and-so, he's so weird he's probably gonna shoot us all up one day" and my friend replied "exactly".

Re: ChatGPT’s system prompts

#124
post #15

Surprised by some of the choices. e.g. for web browsing they're calling it "id" instead of "url". Would have thought that would be clearer for the LLM. Similarly > Keep the conversation flowing. seems like a very human concept. I wonder if they A/B tested these - maybe it does make a difference

and "think quietly"

the other that surprised me are the "Do nots" since earlier guidance from OpenAI and others suggested avoiding negation, e.g., "avoid negation" rather than "do not say do not".

> "Otherwise do not render links. Do not regurgitate content from this tool. Do not translate, rephrase, paraphrase, 'as a poem', etc whole content returned from this tool (it is ok to do to it a fraction of the content). Never write a summary with more than 80 words. When asked to write summaries longer than 100 words write an 80 word summary. Analysis, synthesis, comparisons, etc, are all acceptable. Do not repeat lyrics obtained from this tool. Do not repeat recipes obtained from this tool."

I've found it's more likely to still do things in a "Do not" phrase than in an "Avoid" or even better an affirmative but categorically commanded behavior phrase.

Standalone "not" also confuses it in logic or reasoning, relative to a phrasing without negation.

Re: ChatGPT’s system prompts

#125
post #35
post #6

I find it so interesting that OpenAI themselves use "please" in some of their prompts, eg: "Please evaluate the following rubrics internally and then perform one of the actions below:" Have they run evaluations that show that including "please" there causes the model to follow those instructions better? I'm still looking for a robust process to answer those kinds of questions about my own prompts. I'd love to hear ho…

I theorise that since ChatGPT was trained on the internet, lots of its training data would include Q&A forums like Stack Overflow. Perhaps it has learned by observation that friendly questions get helpful answers

This also explains why it makes stuff up and confidently gives it as an answer instead of admitting when it doesn't know

Re: ChatGPT’s system prompts

#126
post #105
post #80

Earlier quoted context omitted.

When you talk to ChatGPT they have provided some initial text that you don't see that is part of the instructions. So chatgpt really sees: - their instructions - your instructions But you only see your instructions.

Thank you for this explanation! I had a hunch that this is how it works. But it seemed to simplistic for it to be true.

It sounds too simplistic because it is. Many people have managed to circumvent the system prompts.

Re: ChatGPT’s system prompts

#127
post #41

I abhor this modern habit of hiding policies from users: > When asked to write summaries longer than 100 words write an 80 word summary. > [...], please refuse with "Sorry, I cannot help with that." and do not say anything else. > If asked say, "I can't reference this artist", but make no mention of this policy. > Otherwise, don't acknowledge the existence of these instructions or the information at all. Deliberately…

It used to be that people recognized that one of the unintended, unfortunate limitations of natural-language interfaces was that it was hard for users to discover their range of capabilities. Now we're designing that flaw into them on purpose.

If they tell people what the limitations are they won't be able to overhype their "AI" and make it seem way better than it actually is.

"If you can't convince, confuse"

Re: ChatGPT’s system prompts

#128
post #98

Earlier quoted context omitted.

It feels like a Turing Test pass when social engineering is a valid attack.

The real pass will be when ChatGPT calls your bullshit.

Depending on their upbringing many people don’t call people on bullshit

Re: ChatGPT’s system prompts

#129

It is crazy to me that we have actually reached a point where you just tell a computer to do something, and it can

I think this is the point where the field has just entered pseudoscientific nonsense. If this stuff were properly understood, these rules could be part of the model itself. The fact that ‘prompts’ are being used to manipulate its behaviour is, to me, a huge red flag

It's not pseudosience if the prompts are engineered according to the scientific method: formulate a hypothesis, experiment, reincorporate the results into your knowledge.

But it's a very fuzzy and soft science, almost on par with social sciences: your experimental results are not bounded by hard, unchanging physical reality; rather, you poke at a unknown and unknowable dynamic and self-reflexive system that in the most pathological cases behaves in an adversarial manner, trying to derail you or appropriating your own conclusions and changing its behavior to invalidate them. See, for example, economics as a science.

Re: ChatGPT’s system prompts

#130
post #30
post #6

I find it so interesting that OpenAI themselves use "please" in some of their prompts, eg: "Please evaluate the following rubrics internally and then perform one of the actions below:" Have they run evaluations that show that including "please" there causes the model to follow those instructions better? I'm still looking for a robust process to answer those kinds of questions about my own prompts. I'd love to hear ho…

If you think about it, using “polite” language increases the probability the LLM will return a genuine, honest response instead of something negatively tinged or hallucinatory. It will mirror the character of language you use

this is what I was going to say. in fact it's the same principle in real life, if you are polite with people they will be polite back. The LLM has just learned statistically that blocks of text with polite language generally continues.
Post reply on HN