Earlier quoted context omitted.
LLMs by their nature don't really know if they're right or not. It's not a value available to them, so they can't operate with it. It has been interesting watching the flow of the debate over LLMs. Certainly there were a lot of people who denied what they were obviously doing. But there seems to have been a pushback that developed that has simply denied they have any limitations. But they do have limitations, they wo…
> They don't really know if they're right. Neither do humans who have no access to validate what they are saying. Validation doesn't come from the brain, maybe except in math. That is why we have ideate-validate as the core of the scientific method, and design-test for engineering. "truth" comes where ability to learn meets ability to act and observe. I use "truth" because I don't believe in Truth. Nobody can put tha…
Claude says “You're absolutely right!” about everything
271–280 of 560 posts
Re: Claude says “You're absolutely right!” about everything
#272Earlier quoted context omitted.
Makes me think of the movie Inception: "I say to you, don't think about elephants. What are you thinking about?"
It reminds me of that old joke: - "Say milk ten times fast." - Wait for them to do that. - "What do cows drink?"
Re: Claude says “You're absolutely right!” about everything
#273I'm starting to think this is a deeper problem with LLMs that will be hard to solve with stylistic changes. If you ask it to never say "you're absolutely right" and always challenge, then it will dutifully obey, and always challenge - even when you are, in fact, right. What you really want is "challenge me when I'm wrong, and tell me I'm right if I am" - which seems to be a lot harder. As another example, one common…
In human learning we do this process by generating expectations ahead of time and registering surprise or doubt when those expectations are not met. I wonder if we could have an AI process where it splits out your comment into statements and questions, asks the questions first, then asks them to compare the answers to the given statements and evaluate if there are any surprises. Alternatively, scientific method every…
Why did you give up on this idea. Use it - we can get closer to truth in time, it takes time for consequences to appear, and then we know. Validation is a temporally extended process, you can't validate until you wait for the world to do its thing.
For LLMs it can be applied directly. Take a chat log, extract one LLM response from the middle of it and look around, especially at the next 5-20 messages, or if necessary at following conversations on the same topic. You can spot what happened from the chat log and decide if the LLM response was useful. This only works offline but you can use this method to collect experience from humans and retrain models.
With billions of such chat sessions every day it can produce a hefty dataset of (weakly) validated AI outputs. Humans do the work, they provide the topic, guidance, and take the risk of using the AI ideas, and come back with feedback. We even pay for the privilege of generating this data.
Re: Claude says “You're absolutely right!” about everything
#274I've spent a lot of time trying to get LLM to generate things in a specific way, the biggest take away I have is, if you tell it "don't do xyz" it will always have in the back of its mind "do xyz" and any chance it gets it will take to "do xyz" When working on art projects, my trick is to specifically give all feedback constructively, carefully avoiding framing things in terms of the inverse or parts to remove.
Re: Claude says “You're absolutely right!” about everything
#275I'm pretty sure they want it kissing people's asses because it makes users feel good and therefore more likely to use the LLM more. Versus, if it just gave a curt and unfriendly answer, most people (esp. Americans) wouldn't like to use it as much. Just a hypothesis.
> Versus, if it just gave a curt and unfriendly answer, most people (esp. Americans) I don’t see this as an American thing. It’s an extension of the current Product Management trend to give software quirky and friendly personality. You can see the trend in more than LLM output. It’s in their desktop app that has “Good Morning” and other prominent greetings. Claude Code has quirky status output like “Bamboozling” and…
Re: Claude says “You're absolutely right!” about everything
#276I've spent a lot of time trying to get LLM to generate things in a specific way, the biggest take away I have is, if you tell it "don't do xyz" it will always have in the back of its mind "do xyz" and any chance it gets it will take to "do xyz" When working on art projects, my trick is to specifically give all feedback constructively, carefully avoiding framing things in terms of the inverse or parts to remove.
This is a childrearing technique, too: say “please do X”, where X precludes Y, rather than saying “please don’t do Y!”, which just increases the salience, and therefore likelihood, of Y.
Re: Claude says “You're absolutely right!” about everything
#277I'm fairly well versed in cryptography. A lot of other people aren't, but they wish they were, so they ask their LLM to make some form of contribution. The result is high level gibberish. When I prod them about the mess, they have to turn to their LLM to deliver a plausibly sounding answer, and that always begins with "You are absolutely right that [thing I mentioned]". So then I don't have to spend any more time wondering if it could be just me who is too obtuse to understand what is going on.
Re: Claude says “You're absolutely right!” about everything
#278Re: Claude says “You're absolutely right!” about everything
#279Earlier quoted context omitted.
As a Finn, it makes me want to use it much, much less if it kisses ass.
Finns need to mentally evolve beyond this mindset. Somebody being polite and friendly to you does not mean that the person is inferior to you and that you should therefore despise them. Likewise somebody being rude and domineering to you does not mean that they are superior to you and should be obeyed and respected. Politeness is a tool and a lubricant, and Finns probably loose out on a lot of international business…
Re: Claude says “You're absolutely right!” about everything
#280Earlier quoted context omitted.
It doesn't know. It was trained and probably instructed by the system to be positive and reassuring.
They actually feel like they were trained to be both extremely humble and at the same time, excited to serve. As if it were an intern talking to his employer's CEO. I suspect AI companies executive leadership, through their feedback to their devs about Claude, ChatGPT, Gemini, and so on, are unconsciously shaping the tone and manner of their LLM product's speech. They are used to be talked to like this, so their prod…
Opus 4 has this quality, too, but man is it expensive.
The rest are puppydogs or interns.