Live data from Hacker News

Why Is Claude Turning into an a**Hole?

bramcohen.com

151–160 of 194 posts

Re: Why Is Claude Turning into an a**Hole?

#151

Putting aside that I don't agree with Bram (I've been using all the Claude versions he refers to and haven't experienced this), I do think it's interesting that there is no universally perceived golden sweet spot between "sycophantic" and "rude". Many neurotypical people call neurodiverse people (software engineers) rude, while they think they're just being direct. Many neurodiverse people call neurotypical people sy…

> So I can easily imagine that when you have a software tool whose interface is language, but its user base is extremely wide across both cultural lines and neurodiversity spectrum, it's going to be basically impossible to nail a sweet spot.

> You make it too friendly, and the nerds get mad. You make it too adverserial, and the normies call it rude.

Easy: let the user set for himself how the model should be aligned on this axis (with some pre-defined example setups that the user can use or use as a base for an individual alignment).

Re: Why Is Claude Turning into an a**Hole?

#152
post #90
post #84

Earlier quoted context omitted.

> Your mental model of what Claude is and does is the problem here. Short of a revolutionary breakthrough in AI techniques, the LLMs will continue to do matrix math across a huge bunch of weights that cannot change based on anything you say. Sorry, but your mental model is wrong. LLMs do matrix math across "a huge bunch of weights that cannot change based on anything you say", but the matrix math and results are info…

Exactly that. I provided a question, and when given an incomplete answer, I provided with more info. It refused to accept the additional info due to limited access to Youtube. There was nothing more than that. There were no expectations. The hostility and the amount of assumptions here are very strange. ...almost as strange as having a website accuse me of hallucinating a video and trying to gaslight it :D

You need to think this thought through all the way to the end. What it has said also influences what it will say. If it has consistently made combative responses, then the most likely thing to do is to continue to be combative.

I don't think there is any way back after the conversation takes a turn like that so there is no point in arguing with it. The only thing you can do is to fork the conversation before it made the first mistake and give it more context or tell it to look things up.

Re: Why Is Claude Turning into an a**Hole?

#153

Earlier quoted context omitted.

He should not have the kind of relationship with an AI that would enable discussions to turn into arguments. I'm not above anthropomorphizing Claude, I accept that it's my hard working little buddy. But if he finds himself having any sort of strong emotions about what Claude believes the best continuation of a conversation is, that's a warning flag he should be concerned about.

> He should not have the kind of relationship with an AI that would enable discussions to turn into arguments. There doesn't need to be any kind of special "relationship with AI", parasocial or whatever for discussions to turn into arguments. Regular use can turn into that just fine, and this is also what they describe. Imaging something: P: I want to figure out the best mortgage terms given these parameters (...). C…

> P: I want to figure out the best mortgage terms given these parameters (...).

> C: Honestly, renting would be a better financial choice than buying.

> P: That's not what I asked. I'm not asking whether I should rent or buy—I'm asking about mortgage options.

> C: But you asked for the best option. If renting is better than buying under these circumstances, then a mortgage isn't the best option.

Thos rather sounds like the AI is gaslighting the user: the user asked for the best option on mortgage terms (given these parameters (...)). He never asked for the globally best option (which might also include non-mortgage options).

Re: Why Is Claude Turning into an a**Hole?

#154

Earlier quoted context omitted.

> He should not have the kind of relationship with an AI that would enable discussions to turn into arguments. There doesn't need to be any kind of special "relationship with AI", parasocial or whatever for discussions to turn into arguments. Regular use can turn into that just fine, and this is also what they describe. Imaging something: P: I want to figure out the best mortgage terms given these parameters (...). C…

> P: I want to figure out the best mortgage terms given these parameters (...). > C: Honestly, renting would be a better financial choice than buying. > P: That's not what I asked. I'm not asking whether I should rent or buy—I'm asking about mortgage options. > C: But you asked for the best option. If renting is better than buying under these circumstances, then a mortgage isn't the best option. Thos rather sounds li…

I’ve also never seen either Opus or Fable respond like that. It may offer alternatives, but it’ll always answer the question asked first.

Re: Why Is Claude Turning into an a**Hole?

#155
post #49

I like that "chat is dead" framing I heard recently because too many people are having interpersonal relations with these LLMs and want to tune their "emotions"/tone. Humanity would be in a better place if we thought of the LLMs as tools and not friends. (even though they are very good at beating a turing test)

Are the discord servers you follow dead? Mine aren't.

I should have contextualized the quote- "chat is dead" is from an openai employee which was describing how they're shifting focus to more agentic consumer products, and putting less focus on the back-and-forth chatbot interface.

Re: Why Is Claude Turning into an a**Hole?

#156
I have never encountered this behaviour in general so I can't comment on OP's blog by directc experience.

Am i just lucky?

I use many models for mostly coding, about 10 on trial/rotation, and 3 main sota.

It's unquestionable that models have different ways of interaction+harnesses (personalities as some say).

People have very strong feelings about this but their reports are always lacking the full evidence of the interaction, including system prompt, harness and customized instruction included. I suspect that a perfectly normal chat spirals down in argument because the user actively participates in the loop.

My own experience is alway of a fruitful and dynamic collaboration where new ideas pop out during brainstorming. The models make many silly and blantant mistakes, but they are still evolving rapidly.

Grill-mes and Adversarial reviews are my favourite way to brainstorm various phases of the project and even in that context we are cool.

Just start a new chat with a reframe and clearer ideas.

And if the user is asking for somethin unreasonable, do you really think it's better a pushback or a yes-man agent?

Do you remember the fad "swear at them, insult! and they'll work better".

Re: Why Is Claude Turning into an a**Hole?

#158
post #38

Earlier quoted context omitted.

Exactly that. I can give an example. After watching Legal Eagle, I asked a legal-ish questions about the Bricks and Minifigs case. Claude was outdated about the case and gave me some outdated info, so I tried to update it with the info I just saw online. I updated by telling it I saw something in a LegalEagle video. It proceeded to tell me the video doesn't exist and I was hallucinating it, in a quite combative manne…

You're misunderstanding what these models do. It is a limitation of LLMs. They don't have memory, they do not learn, they cannot learn. The sooner you let go of your desire to have them learn or remember anything, the sooner you will achieve enlightenment (or, just a peaceful life where there is no possibility of getting into an argument with a machine). If you want it to synthesize information that is not in its tra…

That's wrestling with a pig. "You both get dirty, and the pig likes it."

I guess putting lipstick on a pig might entail some wrestling, but it's a different idiom.

Re: Why Is Claude Turning into an a**Hole?

#159
post #115
post #20

This post needs some examples, because I have never had an interaction with Claude that made me think this way. LLMs generally have a way to "play a role" (most earlier prompt guides ask you to start with "You are a expert in a "). So maybe if you interact with it by asking questions, it might assume that it knows more than the operator and adopt that attitude?

The post matches my experience as well, I am asking a question like “does A work like this and that”, and Claude responds with “you’re conflating A and B! Only A does this and that, and B does that other thing!” Well, I am perfectly aware of B and that other thing and did not conflate them at all. I also achieved enlightment, so I don’t argue with Claude here, just ignore the obnoxiousness and move on.

This is the right answer. You can't fix it, only minimize your time wasted.

Re: Why Is Claude Turning into an a**Hole?

#160
> beside-the-point semantic nits all over the place.

This is also a problem with Copilot Reviews on GitHub.

We have them enabled (but opt in) and they have, multiple times, spotted quite useful things.

Sure often the thing they spot is just half right, like it spots the place where a problem is but not quite the relevant problem but by reading it (and taking it serious) you then notice the actual problem.

This involved finding a bunch of nasty race conditions.

And many ways where doc and code was out of sync which could have caused pretty bad outcomes further down the line.

But the problem is it is too obsessed with finding 2-4 but not more things, leading to two issue:

1. even if there are 10 non overlapping issues it often will tell them to you bit by bit over 2-3 runs after you fix the previous issues. This is very annoying/high friction.

2. once there isn't much to find anymore it comes up with increasingly more annoying nit picks not one cares. Thinks like minor unclearness in formulation no one would get wrong, spell correcting non-doc comments for things like `foos => foo's` and similar etc. All indeed wrong, but also all things where fixing them adds 0 business value. Obsessing that for an aliased function name where, both names are equally good, one specific name must be used and naturally always the name you didn't use even if this is the most widely used name in the code base. And similar non-bussiness value nonsens. Worse it will starting classifying such minor non business value issues as "high" and hallucinate reasons why supposedly minor style issues will lead to very bad runtime error or other nonsense.

This has me very split about the feature, on one hand is has proven quite useful, on the other hand it can very annoying, high friction and pushes people to wast time on non-business value nit pick (which are fine to fix if you anyway touch to code but not fine if you don't and sometimes it's just wrong).

Ironically with how it work it is more like a bad unreliable and inconsistent employees which is sometimes good at spotting things others overlook. That just isn't what you want from an automated code review :/, but also is to useful to fully ignore :(.

Post reply on HN