Claude's Values Tested in 700K Chats
venturebeat.com
Claude's Values Tested in 700K Chats
1–3 of 3 posts
Re: Claude's Values Tested in 700K Chats
#2Anthropic analyzed 700,000 real conversations to see if Claude behaves the way it was designed. It mostly aligns with their “helpful, honest, harmless” goals — but some edge cases raise big questions.
How should we define and measure values in AI systems — and who decides what they should be?
Re: Claude's Values Tested in 700K Chats
#3[deleted]