Live data from Hacker News

A system prompt to get AI to stop pretending to be human

swiftrocks.com

1–10 of 15 posts

Re: A system prompt to get AI to stop pretending to be human

#5
Do you have a prompt that can stop it from saying, "I'd push back on this"... or maybe with this it will now say "The algorithm pushes back on this"..

Seriously however, I think this may also help with the natural urge to treat the model as if it is a human. I have to purposefully almost detach and realize that Claude is not my friend, and I'm not quite smart enough to realize how dangerous that could be.

Re: A system prompt to get AI to stop pretending to be human

#6
Honestly the models are rled so hard on specific synthetic datasets and specific behaviors/personalities that I would be concerned that trying to change its behavior like this would hurt output quality. It's a tool, I don't care what garbage it generates or what it sounds like as long as it can do what I need it to do, and I don't get what I'm going to gain by having it burn reasoning tokens on word smithing it's responses to not "sound human" instead of on writing tests and reviewing code

Re: A system prompt to get AI to stop pretending to be human

#7
GPT 5.6 seems to have this more than previous versions and more than Claude. It recently said to me "As a PPL holder I would..." So I asked whether it held a PPL :-) the correction was something like "No I'm an AI but a PPL holder would..."

This is probably the result of training on human written Reddit comments that would put it like that.

Re: A system prompt to get AI to stop pretending to be human

#9
I don't have a big issue with writing tonally like a significant portion of its training set (forcing it away from that too hard might not do well). I do have an issue when it literally decides it counts as a human. I have a project where I said "after this step, pause for human review before proceeding". Claude decided it could do the review itself.
Post reply on HN