Live data from Hacker News

Claude's new constitution

anthropic.com

71–80 of 743 posts

Re: Claude's new constitution

#71

The constitution contains 43 instances of the word 'genuine', which is my current favourite marker for telling if text has been written by Claude. To me it seems like Claude has a really hard time _not_ using the g word in any lengthy conversation even if you do all the usual tricks in the prompt - ruling, recommending, threatening, bribing. Claude Code doesn't seem to have the same problem, so I assume the system pr…

You're absolutely right!

[deleted]

Re: Claude's new constitution

#72

I guess this is Anthropic's "don't be evil" moment, but it has about as much (actually much less) weight then when it was Google's motto. There is always an implicit "...for now". No business is every going to maintain any "goodness" for long, especially once shareholders get involved. This is a role for regulation, no matter how Anthropic tries to delay it.

> This is a role for regulation, no matter how Anthropic tries to delay it.

Regulation like SB 53 that Anthropic supported?

https://www.anthropic.com/news/anthropic-is-endorsing-sb-53

Re: Claude's new constitution

#73
LLMs really get in the way of computer security work of any form.

Constantly "I can't do that, Dave" when you're trying to deal with anything sophisticated to do with security.

Because "security bad topic, no no cannot talk about that you must be doing bad things."

Yes I know there's ways around it but that's not the point.

The irony is that LLMs being so paranoid about talking security is that it ultimately helps the bad guys by preventing the good guys from getting good security work done.

Re: Claude's new constitution

#74

So an elaborate version of Asimov's Laws of Robotics? A bit worrying that model safety is approached this way.

One has to wonder, what if a pedophile had an access to nuclear launch codes, and our only hope would be a Claude AI creating some CSAM to distract him from blowing up the world.

But luckily this scenario is already so contrived that it can never happen.

Re: Claude's new constitution

#75
It seems considerably vaguer than a legal document and the verbosity makes it hard to read. I'm tempted to ask Claude for a summary :-)

Perhaps the document's excessive length helps for training?

Re: Claude's new constitution

#76

LLMs really get in the way of computer security work of any form. Constantly "I can't do that, Dave" when you're trying to deal with anything sophisticated to do with security. Because "security bad topic, no no cannot talk about that you must be doing bad things." Yes I know there's ways around it but that's not the point. The irony is that LLMs being so paranoid about talking security is that it ultimately helps th…

Sounds like you need one of them uncensored models. If you don't want to run an LLM locally, or don't have the hardware for it, the only hosted solution I found that actually has uncensored models and isn't all weird about it was Venice. You can ask it some pretty unhinged things.

Re: Claude's new constitution

#77

LLMs really get in the way of computer security work of any form. Constantly "I can't do that, Dave" when you're trying to deal with anything sophisticated to do with security. Because "security bad topic, no no cannot talk about that you must be doing bad things." Yes I know there's ways around it but that's not the point. The irony is that LLMs being so paranoid about talking security is that it ultimately helps th…

I've run into this before too, when playing single player games if I've had enough of grinding sometimes I like to pull up a memory tool, and see if I can increase the amount of wood and so on.

I never really went further but recently I thought it'd be a good time to learn how to make a basic game trainer that would work every time I opened the game but when I was trying to debug my steps, I would often be told off - leading to me having to explain how it's my friends game or similar excuses!

Re: Claude's new constitution

#78

LLMs really get in the way of computer security work of any form. Constantly "I can't do that, Dave" when you're trying to deal with anything sophisticated to do with security. Because "security bad topic, no no cannot talk about that you must be doing bad things." Yes I know there's ways around it but that's not the point. The irony is that LLMs being so paranoid about talking security is that it ultimately helps th…

Last time I tried Codex, it told me it couldn’t use an API token due to a security issue. Claude isn’t too censorious, but ChatGPT is so censored that I stopped using it.

Re: Claude's new constitution

#79
post #72

I guess this is Anthropic's "don't be evil" moment, but it has about as much (actually much less) weight then when it was Google's motto. There is always an implicit "...for now". No business is every going to maintain any "goodness" for long, especially once shareholders get involved. This is a role for regulation, no matter how Anthropic tries to delay it.

> This is a role for regulation, no matter how Anthropic tries to delay it. Regulation like SB 53 that Anthropic supported? https://www.anthropic.com/news/anthropic-is-endorsing-sb-53

Yes, just like that. Supporting regulation at one point in time does not undermine the point that we should not trust corporations to do the right thing without regulation.

I might trust the Anthropic of January 2026 20% more than I trust OpenAI, but I have no reason to trust the Anthropic of 2027 or 2030.

Post reply on HN