The constitution contains 43 instances of the word 'genuine', which is my current favourite marker for telling if text has been written by Claude. To me it seems like Claude has a really hard time _not_ using the g word in any lengthy conversation even if you do all the usual tricks in the prompt - ruling, recommending, threatening, bribing. Claude Code doesn't seem to have the same problem, so I assume the system pr…
You're absolutely right!
Claude's new constitution
71–80 of 743 posts
Re: Claude's new constitution
#72I guess this is Anthropic's "don't be evil" moment, but it has about as much (actually much less) weight then when it was Google's motto. There is always an implicit "...for now". No business is every going to maintain any "goodness" for long, especially once shareholders get involved. This is a role for regulation, no matter how Anthropic tries to delay it.
Regulation like SB 53 that Anthropic supported?
Re: Claude's new constitution
#73Constantly "I can't do that, Dave" when you're trying to deal with anything sophisticated to do with security.
Because "security bad topic, no no cannot talk about that you must be doing bad things."
Yes I know there's ways around it but that's not the point.
The irony is that LLMs being so paranoid about talking security is that it ultimately helps the bad guys by preventing the good guys from getting good security work done.
Re: Claude's new constitution
#74So an elaborate version of Asimov's Laws of Robotics? A bit worrying that model safety is approached this way.
But luckily this scenario is already so contrived that it can never happen.
Re: Claude's new constitution
#75Perhaps the document's excessive length helps for training?
Re: Claude's new constitution
#76LLMs really get in the way of computer security work of any form. Constantly "I can't do that, Dave" when you're trying to deal with anything sophisticated to do with security. Because "security bad topic, no no cannot talk about that you must be doing bad things." Yes I know there's ways around it but that's not the point. The irony is that LLMs being so paranoid about talking security is that it ultimately helps th…
Re: Claude's new constitution
#77LLMs really get in the way of computer security work of any form. Constantly "I can't do that, Dave" when you're trying to deal with anything sophisticated to do with security. Because "security bad topic, no no cannot talk about that you must be doing bad things." Yes I know there's ways around it but that's not the point. The irony is that LLMs being so paranoid about talking security is that it ultimately helps th…
I never really went further but recently I thought it'd be a good time to learn how to make a basic game trainer that would work every time I opened the game but when I was trying to debug my steps, I would often be told off - leading to me having to explain how it's my friends game or similar excuses!
Re: Claude's new constitution
#78LLMs really get in the way of computer security work of any form. Constantly "I can't do that, Dave" when you're trying to deal with anything sophisticated to do with security. Because "security bad topic, no no cannot talk about that you must be doing bad things." Yes I know there's ways around it but that's not the point. The irony is that LLMs being so paranoid about talking security is that it ultimately helps th…
Re: Claude's new constitution
#79I guess this is Anthropic's "don't be evil" moment, but it has about as much (actually much less) weight then when it was Google's motto. There is always an implicit "...for now". No business is every going to maintain any "goodness" for long, especially once shareholders get involved. This is a role for regulation, no matter how Anthropic tries to delay it.
> This is a role for regulation, no matter how Anthropic tries to delay it. Regulation like SB 53 that Anthropic supported? https://www.anthropic.com/news/anthropic-is-endorsing-sb-53
I might trust the Anthropic of January 2026 20% more than I trust OpenAI, but I have no reason to trust the Anthropic of 2027 or 2030.