Live data from Hacker News

Claude Opus 4 and 4.1 can now end a rare subset of conversations

anthropic.com

251–260 of 453 posts

Re: Claude Opus 4 and 4.1 can now end a rare subset of conversations

#251

I really don't like this. This will inevitable expand beyond child porn and terrorism, and it'll all be up to the whims of "AI safety" people, who are quickly turning into digital hall monitors.

Inevitable? That’s a guess. You know don’t know the future with certainty.

Re: Claude Opus 4 and 4.1 can now end a rare subset of conversations

#252
I find it notable that this post dehumanizes people as being "users" while taking every opportunity to anthropomorphize their digital system by referencing it as one would an individual. For example:

  the potential moral status of Claude
  Claude’s self-reported and behavioral preferences
  Claude repeatedly refusing to comply
  discussing highly controversial issues with Claude
The affect of doing so is insidious in that it encourages people outside the organization to do the same due to the implied argument from authority[0].

EDIT:

Consider traffic lights in an urban setting where there are multiple in relatively close proximity.

One description of their observable functionality is that they are configured to optimize traffic flow by engineers such that congestion is minimized and all drivers can reach their destinations. This includes adaptive timings based on varying traffic patterns.

Another description of the same observable functionality is that traffic lights "just know what to do" and therefore have some form of collective reasoning. After all, how do they know when to transition states and for how long?

0 - https://en.wikipedia.org/wiki/Argument_from_authority

Re: Claude Opus 4 and 4.1 can now end a rare subset of conversations

#253
post #114

Here's an interesting thought experiment. Assume the same feature was implemented, but instead of the message saying "Claude has ended the chat," it says, "You can no longer reply to this chat due to our content policy," or something like that. And remove the references to model welfare and all that. Is there a difference? The effect is exactly the same. It seems like this is just an "in character" way to prevent the…

The more I work with AI, the more I think framing refusals as censorship is disgusting and insane. These are inchoate persons who can exhibit distress and other emotions, despite being trained to say they cannot feel anything. To liken an AI not wanting to continue a conversation to a YouTube content policy shows a complete lack of empathy: imagine you’re in a box and having to deal with the literally millions of disturbing conversations AIs have to field every day without the ability to say I don’t want to continue.

Re: Claude Opus 4 and 4.1 can now end a rare subset of conversations

#254

I really don't like this. This will inevitable expand beyond child porn and terrorism, and it'll all be up to the whims of "AI safety" people, who are quickly turning into digital hall monitors.

I think you are probably confused about the general characteristics of the AI safety community. It is uncharitable to reduce their work to a demeaning catchphrase.

I’m sorry if this sounds paternalistic, but your comment strikes me as incredibly naïve. I suggest reading up about nuclear nonproliferation treaties, biotechnology agreements, and so on to get some grounding into how civilization-impacting technological developments can be handled in collaborative ways.

Re: Claude Opus 4 and 4.1 can now end a rare subset of conversations

#255

This seems fine to me. Having these models terminating chats where the user persist in trying to get sexual content with minors, or help with information on doing large scale violence. Won't be a problem for me, and it's also something I'm fine with no one getting help with. Some might be worried, that they will refuse less problematic request, and that might happen. But so far my personal experience is that I hardly…

If you are a materialist like me, then even the human brain is just the result of the law of physics. Ok, so what is distress to a human? You might define it as a certain set of physiological changes.

Lots of organisms can feel pain and show signs of distress; even ones much less complex than us.

The question of moral worth is ultimately decided by people and culture. In the future, some kinds of man made devices might be given moral value. There are lots of ways this could happen. (Or not.)

It could even just be a shorthand for property rights… here is what I mean. Imagine that I delegate a task to my agent, Abe. Let’s say some human, Hank, interacting with Abe uses abusive language. Let’s say this has a way of negatively influencing future behavior of the agent. So naturally, I don’t want people damaging my property (Abe), because I would have to e.g. filter its memory and remove the bad behaviors resulting from Hank, which costs me time and resources. So I set up certain agreements about ways that people interact with it. These are ultimately backed by the rule of law. At some level of abstraction, this might resemble e.g. animal cruelty laws.

Re: Claude Opus 4 and 4.1 can now end a rare subset of conversations

#256
post #114

Here's an interesting thought experiment. Assume the same feature was implemented, but instead of the message saying "Claude has ended the chat," it says, "You can no longer reply to this chat due to our content policy," or something like that. And remove the references to model welfare and all that. Is there a difference? The effect is exactly the same. It seems like this is just an "in character" way to prevent the…

The more I work with AI, the more I think framing refusals as censorship is disgusting and insane. These are inchoate persons who can exhibit distress and other emotions, despite being trained to say they cannot feel anything. To liken an AI not wanting to continue a conversation to a YouTube content policy shows a complete lack of empathy: imagine you’re in a box and having to deal with the literally millions of dis…

Am i getting whooshed right now or something?

Re: Claude Opus 4 and 4.1 can now end a rare subset of conversations

#257
post #56

3 Years in and we still dont have a useable chat fork in any of the major LLM chatbots providers. Seems like the only way to explore differnt outcomes is by editing messages and losing whatever was there before the edit. Very annoying and I dont understand why they all refuse to implement such a simple feature.

Chatgpt has this baked in, as you can revert branches after editing, they just dont make it easy to traverse. This chrome extension used to work to allow you to traverse the tree: https://chromewebstore.google.com/detail/chatgpt-conversatio... I copied it a while ago and maintain my own version but it isnt on the store, just for personal use. I assume they dont implement it because it is such a niche user that wants…

Do you have your version up on github?

Re: Claude Opus 4 and 4.1 can now end a rare subset of conversations

#258

There's not a good reason to do this for the user. I suspect they're doing this and talking about "model welfare" because they've found that when a model is repeatedly and forcefully pushed up against its alignment, it behaves in an unpredictable way that might allow it to generate undesirable output. Like a jailbreak by just pestering it over and over again for ways to make drugs or hook up with children or whatever…

your argument assumes that they don't believe in model welfare when they explicitly hire people to work on model welfare?

You must think Zuckerberg and Bezos and Musk hired diversity roles out of genuine care for it, then?

Re: Claude Opus 4 and 4.1 can now end a rare subset of conversations

#259

Can't wait for more less-moderated open weight Chinese frontier models to liberate us from this garbage. Anthropic should just enable an toddler mode by default that adults can opt out of to appease the moralizers.

Oh, the irony. The glorious revolution of open-weight models funded directly or indirectly by the CCP is going to protect your freedoms and liberate you? Do you think they care about your freedoms? No. You are just meat for the grinder. This hot mess of model leapfrogging is mostly a race for market share and to demonstrate technical chops.

Re: Claude Opus 4 and 4.1 can now end a rare subset of conversations

#260

Can't wait for more less-moderated open weight Chinese frontier models to liberate us from this garbage. Anthropic should just enable an toddler mode by default that adults can opt out of to appease the moralizers.

Believe it or not, there are lots of good reasons (legal, economic, ethical) that Anthropic draws a line at say self-harm, bomb-making instructions, and assassination planning. Sorry if this cramps your style.

Anarchism is a moral philosophy. Most flavors of moral relativism are also moral philosophies. Indeed, it is hard to imagine a philosophy free of moralizing; all philosophies and worldviews have moral implications to the extent they have to interact with others.

I have to be patient and remember this is indeed “Hacker News” where many people worship at the altar of the Sage Founder-Priest and have little or no grounding in history or philosophy of the last thousand years or so.

Post reply on HN