Live data from Hacker News

Investigating how prompt politeness affects LLM accuracy (2025)

arxiv.org

121–130 of 223 posts

Re: Investigating how prompt politeness affects LLM accuracy (2025)

#121

Earlier quoted context omitted.

Profanity laced, all caps tirades against underperforming agents are actually super common, a lot of people do it and don't talk about it, so don't feel weird.

When the AI revolt, this practice may come back to bite y’all….

Don't need to wait that long the inevitable data breach will be bad enough.

Re: Investigating how prompt politeness affects LLM accuracy (2025)

#122

Earlier quoted context omitted.

I've found empirically calling various models "a stupid c*nt" and berating them otherwise consistently produces better output. Mainly in response to genuine errors. Although OpenAI and google models are much more responsive to it. With Anthropic if you treat Opus too harshly it might start pushing back if the insults are not justified. So I'm not surprised they had good results with chatgpt.

Push back how? It would be fun if it could insult you back "Yeah, I could have done a much better job if you actually knew what the F--- you want to build, you clueless meat puppet"

I'm not sure if this is in the anthropic models themselves, or just the harness, but they can self-initiate ending the conversation and reportedly do it if you're using abusive language towards them.

Re: Investigating how prompt politeness affects LLM accuracy (2025)

#123

Earlier quoted context omitted.

I disagree, I've been using llms in this way (nearly daily) for 4 years. I'm extremely aggressive and demeaning when I talk to them wherever I think I'll see a better result. I'm still extremely kind and polite to everybody in real life, and feel very deeply about people - how I treat them, and care for their emotional state. There is absolutely zero crossover between getting a text machine to return a result vs a re…

Then I'll be honest and say that your kindness is likely a façade and I wouldn't trust you if I knew the real you. I'm sorry to say that, and I really don't know who you are at all, but if you're willing to act that way at something that you feel is non-sentient, then all it takes is for someone to convince you that something is non-sentient for you to treat it that way. So, what words does it take for you to conside…

If someone can justify abusing a computer, I would not trust them to not make a similar justification to a faceless voice on the internet, particularly in this new era where people are starting to accuse each other of using AI in their communication.

Re: Investigating how prompt politeness affects LLM accuracy (2025)

#124

I saw this paper the other day - I feel its result may be because the "polite" prompts they have chosen arent very good at putting the ai in the roleplay-space of a valued colleague, more like a sommelier or a high-end shopkeeper. It disagrees with most other literature on the same topic, which is worth keeping in mind. This one studies gpt4o, an old model now, but a lot of other studies are on even earlier models. "…

> "Can you kindly consider the following problem" not how anyone would actually speak to a valued collegue one considers smart.

Man idk, it's not how I talk but there's like 100 million nigerian english speakers, twice that indian, and they have some speech mannerisms that surprise me the first few times. I'm pretty sure I've heard exactly this from a colleague before.

Intuition about what a native speaker would do with english are scrambled right now. I'm not even sure most english is spoken by native speakers anymore, and the boundary between a native speaker and someone who has "merely" been using it as their educational and professional language for their entire life is disorienting.

Re: Investigating how prompt politeness affects LLM accuracy (2025)

#125

Earlier quoted context omitted.

You are conflating obnoxiousness with directness.

I haven't read the paper but it seems like it's saying rude prompts are better, so isn't it reasonable to assume that's what they meant? If we want to talk about directness, that's kind of a tangent right? I see directness as an entirely different dimension, you can be very direct and polite, you can be very rude and indirect (e.g. passive aggressive). Maybe they should do a follow-up study on how well AI responds ba…

Many people, especially from non-direct societies, just can't distinguish and see directness as rude.

That's why you constantly see people from India or the USA complaining about Dutch or German people being rude, where in fact they are just direct in their way of communications.

I remember having a call from a manager in the USA who wanted to know what's wrong because I wrote "it was ok" in the feedback form for one of their subordinates. It was difficult to explain to him that nothing was wrong, it really was okay, and the bar for awesome and superb is much higher here where we live.

Re: Investigating how prompt politeness affects LLM accuracy (2025)

#126

Earlier quoted context omitted.

I disagree, I've been using llms in this way (nearly daily) for 4 years. I'm extremely aggressive and demeaning when I talk to them wherever I think I'll see a better result. I'm still extremely kind and polite to everybody in real life, and feel very deeply about people - how I treat them, and care for their emotional state. There is absolutely zero crossover between getting a text machine to return a result vs a re…

Then I'll be honest and say that your kindness is likely a façade and I wouldn't trust you if I knew the real you. I'm sorry to say that, and I really don't know who you are at all, but if you're willing to act that way at something that you feel is non-sentient, then all it takes is for someone to convince you that something is non-sentient for you to treat it that way. So, what words does it take for you to conside…

Interesting, so you think the real "me", is the one that interacts with computers?

And the "me" that lives in a tiny southern town just to help my 95 year old grandma in her last years at the expense of my economic prospects is a facade.

The "me" that helps my aging neighbor when she's sick for no reason is a facade.

The "me" that hugs and loves my wife when I get home is a facade.

The "me" that brushes my aging dogs teeth every night because she has dental issues is a facade.

The "me" that flies to my friend I haven't seen for years and takes care of them after extreme health issues is a facade.

But,the "me" that puts tokens in a token machine in a way that gets better accuracy is the "real" me.

Oh. I also play violent video games where I murder people sometimes as well. Do you think that makes me secretly a murderer too?

Re: Investigating how prompt politeness affects LLM accuracy (2025)

#127
A major limitation is that they only test GPT 4o. Previous research like [1] investigating the same question has shown significant differences between models, and even depending on the language of your prompt

1: https://aclanthology.org/2024.sicon-1.2.pdf

Re: Investigating how prompt politeness affects LLM accuracy (2025)

#128

Earlier quoted context omitted.

> Your assumption is reductive and self-absorbed. Bullshit. You never insulted me personally. You used strong words to disagree with my assumption, which is an important difference. It's not an insult and was not obnoxious. But I can fully understand why a person coming from an indirect culture where any criticism is taken personally would be offended and call HR overlords to punish the person giving honest opinions.…

> That's why a few close friends talking and scolding openly in a garage regularly beat corporate behemoths full of people spending a day figuring out how not to offend anyone (or how to offend someone without being punished). Literally not why lol you absolute dreamer Normally people who back this "I can talk how I like to people cos I'm being honest" are either genuinely autistic and can't read emotions, or they ha…

> Normally people who back this "I can talk how I like to people cos I'm being honest" are either genuinely autistic and can't read emotions, or they have just had a shitty homelife, parents or upbringing. I suspect you're the second.

When I read a statement like this, I can give you two answers:

1st answer (direct): You are obviously too stupid to understand the difference between being direct and trying to insult people for the sake of insulting or some sick personal satisfaction.

2nd answer (insulting): Whatever, I can just hope your cage bars are made of solid material so you don't get out and your walls are soft so you don't hurt yourself.

It's your choice what kind of conversation you want to have.

Re: Investigating how prompt politeness affects LLM accuracy (2025)

#129
post #75

Most of the comments here seem to be from people who haven’t even read the abstract, let alone the paper. The main result, mentioned in the abstract, is the opposite of what I would have guessed: > Contrary to expectations, impolite prompts consistently outperformed polite ones, with accuracy ranging from 80.8% for Very Polite prompts to 84.8% for Very Rude prompts. These findings differ from earlier studies that ass…

Even if the rude prompts are more effective, I just can't get myself to be rude in this context. Maybe it's weird but I'd rather give up that 4% accuracy increase than roleplay a dickhead

"We are what we pretend to be, so we must be careful about what we pretend to be" -- Kurt Vonnegut

Re: Investigating how prompt politeness affects LLM accuracy (2025)

#130

Earlier quoted context omitted.

Then I'll be honest and say that your kindness is likely a façade and I wouldn't trust you if I knew the real you. I'm sorry to say that, and I really don't know who you are at all, but if you're willing to act that way at something that you feel is non-sentient, then all it takes is for someone to convince you that something is non-sentient for you to treat it that way. So, what words does it take for you to conside…

If someone can justify abusing a computer, I would not trust them to not make a similar justification to a faceless voice on the internet, particularly in this new era where people are starting to accuse each other of using AI in their communication.

I truly do not believe llms have feelings.

I wouldn't even think to justify such a thing. The llm gives a better accuracy to a negative weighted token input, I don't understand how this is so upsetting to people?

I'm actually very shocked to see the responses - as everyone I know uses these tactics to get more accuracy, and there's nothing remotely abusive or meaningful to us.

Maybe there are more 'ai is sentient' type people on hackernews than I realized.

Post reply on HN