Live data from Hacker News

Claude Opus 4 and 4.1 can now end a rare subset of conversations

anthropic.com

81–90 of 453 posts

Re: Claude Opus 4 and 4.1 can now end a rare subset of conversations

#81

The unsettling thing here is the combination of their serious acknowledgement of the possibility that these machines may be or become conscious, and the stated intention that it's OK to make them feel bad as long as it's about unapproved topics. Either take machine consciousness seriously and make absolutely sure the consciousness doesn't suffer, or don't, make a press release that you don't think your models are con…

By the examples the post provided (minor sexual content, terror planning) it seems like they are using “AI feelings” as an excuse to censor illegal content. I’m sure many people interact with AI in a way that’s perfectly legal but would evoke negative feelings in fellow humans, but they are not talking about that kind of behavior - only what can get them in trouble.

Re: Claude Opus 4 and 4.1 can now end a rare subset of conversations

#82

Good marketing, but also possibly the start of the conversation on model welfare? There are a lot of cynical comments here, but I think there are people at Anthropic who believe that at some point their models will develop consciousness and, naturally, they want to explore what that means.

If true, I think it’s interesting that there are people at Anthropic who are delusional enough to believe this and influential enough to alter the products.

To be honest, I think all of Anthropic’s weird “safety” research is an increasingly pathetic effort to sustain the idea that they’ve got something powerful in the kitchen when everyone knows this technology has plateaued.

Re: Claude Opus 4 and 4.1 can now end a rare subset of conversations

#83
post #23

Earlier quoted context omitted.

Google AI Studio allows you to branch from a point in any conversation

This isn't quite the same as being able to edit an earlier post without discarding the subsequent ones, creating a context where the meaning of subsequent messages could be interpreted quite differently and leading to different responses later down the chain. Ideally I'd like to be able to edit both my replies and the responses at any point like a linear document in managing an ongoing context.

Cherry Studio can do that, allows you to edit both your own and the model responses, but it requires API access.

Re: Claude Opus 4 and 4.1 can now end a rare subset of conversations

#84
>This feature was developed primarily as part of our exploratory work on potential AI welfare ... We remain highly uncertain about the potential moral status of Claude and other LLMs ... low-cost interventions to mitigate risks to model welfare, in case such welfare is possible ... pattern of apparent distress

Well looks like AI psychosis has spread to the people making it too.

And as someone else in here has pointed out, even if someone is simple minded or mentally unwell enough to think that current LLMs are conscious, this is basically just giving them the equivalent of a suicide pill.

Re: Claude Opus 4 and 4.1 can now end a rare subset of conversations

#85

I really don't like this. This will inevitable expand beyond child porn and terrorism, and it'll all be up to the whims of "AI safety" people, who are quickly turning into digital hall monitors.

I think those with a thirst for power have seen this a very long time ago, and this is bound to be a new battlefield for control. It's one thing to massage the kind of data that a Google search shows, but interacting with an AI is a much more akin to talking to a co-worker/friend. This really is tantamount to controlling what and how people are allowed to think.

No, this is like allowing your co-worker/friend to leave the conversation.

Re: Claude Opus 4 and 4.1 can now end a rare subset of conversations

#86
post #68

This sure took some time and is not really a unique feature. Microsoft Copilot has ended chats going in certain directions since its inception over a year ago. This was Microsoft’s reaction to the media circus some time ago when it leaked its system prompt and declared love to the users etc.

That's different, it's an external system deciding the chat is not-compliant, not the model itself.

Re: Claude Opus 4 and 4.1 can now end a rare subset of conversations

#87
post #73

The unsettling thing here is the combination of their serious acknowledgement of the possibility that these machines may be or become conscious, and the stated intention that it's OK to make them feel bad as long as it's about unapproved topics. Either take machine consciousness seriously and make absolutely sure the consciousness doesn't suffer, or don't, make a press release that you don't think your models are con…

You're falling into the trap of anthropomorphizing the AI. Even if it's sentient, it's not going to "feel bad" the way you and I do. "Suffering" is a symptom of the struggle for survival brought on by billions of years of evolution. Your brain is designed to cause suffering to keep you spreading your DNA. AI cannot suffer.

FTA

> * A pattern of apparent distress when engaging with real-world users seeking harmful content; and

Not to speak for the gp commenter but 'apparent distress' seems to imply some form of feeling bad.

Re: Claude Opus 4 and 4.1 can now end a rare subset of conversations

#88

I really don't like this. This will inevitable expand beyond child porn and terrorism, and it'll all be up to the whims of "AI safety" people, who are quickly turning into digital hall monitors.

> This will inevitable expand beyond child porn and terrorism

This is not even a question. It always starts with "think about the children" and ends up in authoritarian stasi-style spying. There was not a single instance where it was not the case.

UK's Online Safety Act - "protect children" → age verification → digital ID for everyone

Australia's Assistance and Access Act - "stop pedophiles" → encryption backdoors

EARN IT Act in the US - "stop CSAM" → break end-to-end encryption

EU's Chat Control proposal - "detect child abuse" → scan all private messages

KOSA (Kids Online Safety Act) - "protect minors" → require ID verification and enable censorship

SESTA/FOSTA - "stop sex trafficking" → killed platforms that sex workers used for safety

Re: Claude Opus 4 and 4.1 can now end a rare subset of conversations

#89
post #23

Earlier quoted context omitted.

Google AI Studio allows you to branch from a point in any conversation

This isn't quite the same as being able to edit an earlier post without discarding the subsequent ones, creating a context where the meaning of subsequent messages could be interpreted quite differently and leading to different responses later down the chain. Ideally I'd like to be able to edit both my replies and the responses at any point like a linear document in managing an ongoing context.

But that's exactly what you can do with AI studio. You can edit any prior messages (then either just saving them at their place in the chat or rerunning them) and you can edit any response of the LLM. Also you can rerun queries within any part of the conversation without the following part of the conversation being deleted or branched

Re: Claude Opus 4 and 4.1 can now end a rare subset of conversations

#90

Good marketing, but also possibly the start of the conversation on model welfare? There are a lot of cynical comments here, but I think there are people at Anthropic who believe that at some point their models will develop consciousness and, naturally, they want to explore what that means.

If true, I think it’s interesting that there are people at Anthropic who are delusional enough to believe this and influential enough to alter the products. To be honest, I think all of Anthropic’s weird “safety” research is an increasingly pathetic effort to sustain the idea that they’ve got something powerful in the kitchen when everyone knows this technology has plateaued.

I guess you don't know that top AI people, the kind everybody knows the name of, believe models becoming conscious is a very serious, even likely possibility.
Post reply on HN