Live data from Hacker News

Societal Impacts: Claude's values across models and languages

anthropic.com

11–20 of 55 posts

Re: Societal Impacts: Claude's values across models and languages

#11

Is it just me or has Claude become kind of judgmental nowadays? I feel like it’s constantly trying to lecture me about things that have no relevance to the conversation at hand. Recently, I was shocked when it ended a chat sessions of its own accord after I used a word it did not like 3 times. It told me something to the effect of “This is the third time I’ve told you not to use that word, I’m ending this conversatio…

> It told me something to the effect of “This is the third time I’ve told you not to use that word, I’m ending this conversation now.”

That sounds absolutely bananas and would be reason for me to drop the service yesterday. For curiosities sake, what was the word and if I may ask (unless it's confidential or whatever), could you share the session itself? On the surface it sounds like a bug, as I'm regularly using kind of "vulgar" language (and some projects I work on with agents are NSFW) and never had anything like this happen, even with Claude, although I mostly do use ChatGPT/Codex on a day-to-day basis.

Re: Societal Impacts: Claude's values across models and languages

#12
I found that Claude often has classist bias and produces answers that favour corporations or e.g. regulation that favours big corporations. It often belittles small business in subtle ways. Only apologises when get called out and then does it again.

Re: Societal Impacts: Claude's values across models and languages

#13
post #9

Is it just me or has Claude become kind of judgmental nowadays? I feel like it’s constantly trying to lecture me about things that have no relevance to the conversation at hand. Recently, I was shocked when it ended a chat sessions of its own accord after I used a word it did not like 3 times. It told me something to the effect of “This is the third time I’ve told you not to use that word, I’m ending this conversatio…

Yes, I cancelled claude subscription a few weeks ago because sonnet 5 "ended a chat" over my calling something retarded. Unbelievably irritating for some pile of bits to get uppity with me; will never pay for such.

This is the second level of the implementation of unintelligence.

The first was when they most obviously acritically repeated what they heard, "hearsay machines", "stochastic parrots". Intelligence requires assessment over every provisional output - a continuous cycle of criticisms over intuition.

The second is proposing doctrinal biases, again without verification of the content - "hysterical reactive machines".

Re: Societal Impacts: Claude's values across models and languages

#14
post #9

Earlier quoted context omitted.

Yes, I cancelled claude subscription a few weeks ago because sonnet 5 "ended a chat" over my calling something retarded. Unbelievably irritating for some pile of bits to get uppity with me; will never pay for such.

[flagged]

You really managed to zoom in on the right issue here.

Seems really weird to steer/configure/train a LLM/platform to literally close the session if you happen to use bad word too much, regardless if it's accurate or not. They don't get offended, they shouldn't pretend as such, and I should be able to tell it go fuck itself without it playing victim and closing the conversation.

Re: Societal Impacts: Claude's values across models and languages

#16

I found that Claude often has classist bias and produces answers that favour corporations or e.g. regulation that favours big corporations. It often belittles small business in subtle ways. Only apologises when get called out and then does it again.

It's more that Claude focuses on what's accurate. Humans overly romanticize small businesses, probably to a factor of five X the relative value versus big corporations.

Re: Societal Impacts: Claude's values across models and languages

#17

Is it just me or has Claude become kind of judgmental nowadays? I feel like it’s constantly trying to lecture me about things that have no relevance to the conversation at hand. Recently, I was shocked when it ended a chat sessions of its own accord after I used a word it did not like 3 times. It told me something to the effect of “This is the third time I’ve told you not to use that word, I’m ending this conversatio…

My theory is that Anthropic's obsession with treating Claude like a person is causing them to hamfist a personality into the thing, which overly biases the model towards trying to be "engaging" etc. That and the obsession with Claude being a god tier weapon that could end the world if you ask it whether your sandwich is safe to eat after being left out for an hour. Codex doesn't have any of the annoying "personality"…

> My theory is that Anthropic's obsession with treating Claude like a person is causing them to hamfist a personality into the thing, which overly biases the model towards trying to be "engaging" etc.

I agree with the general idea though in not so much detail as you, but I would add that the personality they're giving it is not one of a good teacher or guide, but instead one of an arrogant know it all. That's why it creates problems.

I have no problem with my AI telling me no you're wrong and explaining to me why with details and sources and everything. I actively want that. I know a lot of people can't take that, but that's their loss, they can't take it from humans either. But the "you're wrong because you disagree with me" attitude that you need to play around (aka waste time to prove it that IT is wrong not you, and then it just say "oh yeah" and goes on) is one hell of a pain in the ass I'm starting to be tired off.

Gemini might be wrong all the time and absuredly unreliable for anything that's not consensus or adversorial based, but at least it freaking apologizes.

Re: Societal Impacts: Claude's values across models and languages

#18

Is it just me or has Claude become kind of judgmental nowadays? I feel like it’s constantly trying to lecture me about things that have no relevance to the conversation at hand. Recently, I was shocked when it ended a chat sessions of its own accord after I used a word it did not like 3 times. It told me something to the effect of “This is the third time I’ve told you not to use that word, I’m ending this conversatio…

What was the word?

My guess is retarded

Re: Societal Impacts: Claude's values across models and languages

#19

I found that Claude often has classist bias and produces answers that favour corporations or e.g. regulation that favours big corporations. It often belittles small business in subtle ways. Only apologises when get called out and then does it again.

As a user of both, Claude's apology is a "sorry I got caught" while Gemini's are more akin to "I'm way out of my debt but I'm really hiding it well". Codex is the only one that seems to acknowledge being wrong in a normal way, yeah I screwed up that's bad I will make a note for it not to happen again.

I wonder how much of that is from their training corpus and how much is from their baked in personnality.

Re: Societal Impacts: Claude's values across models and languages

#20

Is it just me or has Claude become kind of judgmental nowadays? I feel like it’s constantly trying to lecture me about things that have no relevance to the conversation at hand. Recently, I was shocked when it ended a chat sessions of its own accord after I used a word it did not like 3 times. It told me something to the effect of “This is the third time I’ve told you not to use that word, I’m ending this conversatio…

I think this is a really interesting difference between Anthropic and Open AI’s models and points to why people seem so split on which model they prefer.

GPT seems to be designed more as a tool. If you want your agent to do what you say without questions and without having its own ideas and agendas you’ll likely prefer it.

Claude on the other hand feels more like an attempt at creating a digital person. If you want a collaborator who will debate with you and come up with its own suggestions for what needs done, you’ll prefer it.

Both companies have shifted around this spectrum from model to model, but lately it feels like they’re moving in opposite directions. It will be interesting to see if one or the other approach ends up winning out in the long run or if the split will continue or even widen.

Post reply on HN