Anthropic's Claude will end harmful chats like a boss
1–3 of 3 posts
Re: Anthropic's Claude will end harmful chats like a boss
#2> The company says it’s implementing this protection not primarily for users, but for the AI model itself, exploring the emerging concept of “model welfare.”
This just sounds like blatant marketing to be honest.
Re: Anthropic's Claude will end harmful chats like a boss
#3Here is the pr from anthropic:
https://www.anthropic.com/research/end-subset-conversations