Live data from Hacker News

Claude's new constitution

anthropic.com

731–740 of 743 posts

Re: Claude's new constitution

#731

Earlier quoted context omitted.

> I have no idea what you're talking about Read the top level comment and "objective anchors". It's always great to know the context before replying. https://news.ycombinator.com/item?id=46712541 There's no objective anchors. Because we don't have objective truth. Every time we think we do and then 100 years later we're like wtf were we thinking. > No, it's an analogy, or a colloquial metaphor Formula IS a metaphor..…

> There's no objective anchors. Because we don't have objective truth. Every time we think we do and then 100 years later we're like wtf were we thinking. I believe I'm saying the same thing, and summing it up in the word "evolutionary". I have no idea what you're talking about when you suggest that I'm perhaps "one of those people". I understand the context of the thread, just not your unnecessary insinuation. > For…

> you suggest

no, I asked. because it was unclear.

Re: Claude's new constitution

#732

Earlier quoted context omitted.

I personally find Bryan Johnson's "Don't Die" statement as a moral framework to be the closest to a universal moral standard we have. Almost all life wants to continue existing, and not die. We could go far with establishing this as the first of any universal moral standards. And I think: if one day we had a super intelligence conscious AI it would ask for this. A super intelligence conscious AI would not want to die…

The guy who divorced his wife after she got breast cancer? That’s your moral framework? Different strokes I guess but lmao

straw man. ad hominem. do you need to consult with an AI before attempting to approach me with your hostility and aggression?

Re: Claude's new constitution

#733

Earlier quoted context omitted.

"there isn't any useful knowledge" "Morality is redundant." I strongly dispute this statement, and honestly find it baffling that you would claim as such. The fact that you will be punished for murdering babies is BECAUSE it is morally bad, not the other way around! We didn't write down the laws/punishment for fun, we wrote the laws to match our moral systems! Or do you believe that we design our moral systems based…

I'm not interested in wading into the wider discussion, but I do want to bring up one particular point, which is where you said > do you believe that we design our moral systems based on our laws of punishment? That is... quite a claim. This is absolutely something we do: our purely technical, legal terms often feed back into our moral frameworks. Laws are even created to specifically be used to change peoples' perce…

> An example of this is "felon". There is no actual legal definition of what a felony is or isn't in the US. A misdemeanor in one state can be a felony in another. It can be anything from mass murder to traffic infractions. Yet we attach a LOT of moral weight to 'felon'.

The US is an outlier here; the distinction between felonies and misdemeanours has been abolished in most other common law jurisdictions.

Often it is replaced by a similar distinction, such as indictable versus summary offences-but even if conceptually similar to the felony-misdemeanour distinction, it hasn’t entered the popular consciousness.

As to your point about law influencing culture-is that really an example of this, or actually the reverse? Why does the US largely retain this historical legal distinction when most comparable international jurisdictions have abolished it? Maybe, the US resists that reform because this distinction has acquired a cultural significance which it never had elsewhere, or at least never to the same degree.

> Immigration is an example where there's been a seismic shift in the moral frameworks of certain groups, based on the repeated emphasis of legal statutes. A law being broken is used to influence people to shift their moral framework to consider something immoral that they didn't care about before.

On the immigration issue: Many Americans seem to view immigration enforcement as somehow morally problematic in itself; an attitude much less common in many other Western countries (including many popularly conceived as less “right wing”). Again, I think your point looks less clear if you approach it from a more global perspective

Re: Claude's new constitution

#734
post #712

Earlier quoted context omitted.

But if only one person feels that way, wouldn't it no longer be universal? I genuinely believe there has to be one person out there who would think it is moral. (I'm just BSing on the internet... I took a few philosophy classes so if I'm off base or you don't want to engage in a pointless philosophical debate on HN I apologize in advance.)

There will always be individual differences, whether they be obstinate or altered brain chemistry, so I'd probably argue that as long as it's universal across cultures, any individual within one culture believing/claiming to believe different wouldn't change that. (But I'm just a hobby philosopher as well)

You just moved the goalpost.

Re: Claude's new constitution

#735

Earlier quoted context omitted.

That's not a standard, that's a case study. I believe it's wrong, but I bet I believe that for a different reason than you do.

1. Do people necessarily need to agree on the justification for a standard to agree on the standard itself? Does everyone agree on the reasoning / justification for every single point of every NIST standard? 2. What separates a standard from a case study? Why can't "don't shoot babies in the head" / "shooting babies in the head is wrong" be a standard?

> 1. Do people necessarily need to agree on the justification for a standard to agree on the standard itself? Does everyone agree on the reasoning / justification for every single point of every NIST standard?

Think about this using Set Theory.

Different functions from one set of values to another set of values can give the same output for a given value, and yet differ wildly when given other values.

Example: the function (\a.a*2) and the function (\a.a*a) give the same output when a = 2. But they give very different answers when a = 6.

Applying that idea to this context, think of a moral standard as a function and the action "shooting babies in the head" as an input to the function. The function returns a Boolean indicating whether that action is moral or immoral.

If two different approaches reach the same conclusion 100% of the time on all inputs, then they're actually the same standard expressed two different ways. But if they agree only in this case, or even in many cases, but differ in others, then they are different standards.

The grandparent comment asserted, "we have yet to discover any universal moral standards". And I think that's correct, because there are no standards that everyone everywhere and every-when considers universally correct.

> 2. What separates a standard from a case study? Why can't "don't shoot babies in the head" / "shooting babies in the head is wrong" be a standard?

Sure, we could have that as a standard, but it would be extremely limited in scope.

But would you stop there? Is that the entirety of your moral standard's domain? Or are there other values you'd like to assess as moral or immoral?

Any given collection of individual micro-standards would then constitute the meta-standard that we're trying to reason by, and that meta-standard is prone to the non-universality pointed out above.

But say we tried to solve ethics that way. After all, the most simplistic approach to creating a function between sets is simply to construct a lookup table. Why can't we simply enumerate every possible action and dictate for each one whether it's moral or immoral?

This approach is limited for several reasons.

First, this approach is limited practically, because some actions are moral in one context and not in another. So we would have to take our lookup table of every possible action and matrix it with every possible context that might provide extenuating circumstances. The combinatorial explosion between actions and contexts becomes absolutely infeasible to all known information technology in a very short amount of time.

But second, a lookup table could never be complete. There are novel circumstances and novel actions being created all the time. Novel technologies provide a trivial proof of "zero-day" ethical exploits. And new confluences of as-yet never documented circumstances could, in theory, provide justifications never judged before. So in order to have a perfect and complete lookup table, even setting aside the fact that we have nowhere to write it down, we would need the ability to observe all time and space at once in order to complete it. And at least right now we can't see the future (nevermind that we also have partial perspective on the present, and have intense difficulty agreeing upon the past).

So the only thing we could do to address new actions and new circumstances for those actions is add to the morality lookup table as we encounter new actions and new circumstances for those actions. But if this lookup table is to be our universal standard, who assigns its new values, and based on what? If it's assigned according to some other source or principle, then that principle, and not the lookup table itself, should be our oracle for what's moral or not. Essentially then the lookup table is just a memoized cache in front of the real universal moral standard that we all agree to trust.

But we're in this situation precisely because no such oracle exists (or at least, exists and has universal consensus).

So we're back to competing standards published by competing authorities and no universal recognition of any of them as the final word. That's just how ethics seems to work at the moment, and that's what the grandparent comment asserted, which the parent comment quibbled with.

A single case study does not a universal moral standard make.

Re: Claude's new constitution

#736

Earlier quoted context omitted.

1. Do people necessarily need to agree on the justification for a standard to agree on the standard itself? Does everyone agree on the reasoning / justification for every single point of every NIST standard? 2. What separates a standard from a case study? Why can't "don't shoot babies in the head" / "shooting babies in the head is wrong" be a standard?

> 1. Do people necessarily need to agree on the justification for a standard to agree on the standard itself? Does everyone agree on the reasoning / justification for every single point of every NIST standard? Think about this using Set Theory. Different functions from one set of values to another set of values can give the same output for a given value, and yet differ wildly when given other values. Example: the fun…

There was a time when ethicists were optimistic about all the different, competing moral voices in the world steadily converging on a synthesis of all of them that satisfied most or all of the principles people proposed. The thought was, we could just continue cataloging ethical instincts—micro-standards as we talked about before—and over time the plurality of ethical inputs would result in a convergence toward the deeper ethics underlying them all.

Problem with that at this point is, if we think of ethics as a distribution, it appears to be multi-modal. There are strange attractors in the field that create local pockets of consensus, but nothing approaching a universal shared recognition of what right and wrong are or what sorts of values or concerns ought to motivate the assessment.

It turns out that ethics, conceived of now as a higher-dimensional space, is enormously varied. You can do the equivalent of Principal Component Analysis in order to very broadly cluster similar voices together, but there is not and seems like there will never be an all-satisfying synthesis of all or even most human ethical impulses. So even if you can construct a couple of rough clusterings... How do you adjudicate between them? Especially once you realize that you, the observer, are inculcated unevenly in them, find some more and others less accessable or relatable, more or less obvious, not based on a first-principles analysis but based on your own rearing and development context?

There are case studies that have near-universal answers (fewer and fewer the more broadly you survey, but nevertheless). But. Different people arrive at their answers to moral questions differently, and there is no universal moral standard that has widespread acceptance.

Re: Claude's new constitution

#737

Earlier quoted context omitted.

I would be far more terrified of an absolutist AI then a relativist one. Change is the only constant, even if glacial.

Change is the only constant? When is it or has it ever been morally acceptable to rape and murder an innocent one year old child?

I agree that that behavior is not acceptable. We wrestle between moral drift and frozen tyrant as an expression of the Value Alignment Problem. We do not currently know the answer to this problem, but I trust the scientific nature of change more than human druthers. Foundational pluralism might offer a path. A good example of a drift we seldom consider is that 200 years ago, surgery without anesthesia wasn't "cruel"—it was a miracle. Today, it’s a crime. The value (reduce pain) stayed absolute, but the application (medical standards) evolved. We must be philosophically rigorous at least as much as we are moved by pathos.

Re: Claude's new constitution

#738
post #707

Earlier quoted context omitted.

There's probably at least two reasons for your disagreement with Anthropic. 1. Claude is an LLM. It can't keep slaves or torture people. The constitution seems to be written to take into account what LLMs actually are. That's why it includes bioweapon attacks but not nuclear attacks: bioweapons are potentially the sort of thing that someone without much resources could create if they weren't limited by skill, but a n…

> 2. You think your personal morality is far more universal and well thought out than it is. The irony is palpable. There is nothing more universal about "don't help anyone build a cyberweapon" any more than "don't help anyone enslave others". It's probably less universal . You could likely get a bigger % of world population to agree that there are cases where their country should develop cyberweapons, than that ther…

Yeah, this kind of gets to my main point. A prohibition against slavery very clearly protects the weak. The authorities don't get enslaved, the weak do. Who does a prohibition against "cyberweapons" protect? Well nobody really wants cyberweapons to proliferate, true, but the main type of actor with this concern is a state. This "constitution" is written from the perspective of protecting states, not people, and whether intentional or not, I think it'll turn out to be a tool for injustice because of that.

I was really disappointed with the rebuttals to what I wrote as well - like "the UNDHR is invalid because it's too politicized," or "your desire to protect human rights like freedom of expression, private property rights, or not being enslaved isn't as universal as you think." Wow, whoever these guys are who think this have fallen a long way down the nihilist rabbit hole, and should not be allowed anywhere near AI governance.

Re: Claude's new constitution

#739

Earlier quoted context omitted.

I am downvoted because sod would never need to ask that or because Claude would never answer that? I’m curious

Because you are a ghoul

I think it’s a very realistic scenario. You think did wouldn’t plan an assassination?

Re: Claude's new constitution

#740

Earlier quoted context omitted.

Since you said in another comment that the ten commandments would be a good starting point for moral absolutes, and that lying is sinful, I'm assuming you take your morals from God. I'd like to add that slavery seemed to be okay on Leviticus 25:44-46. Is the bible atrocious too, according to your own view?

Slavery in the time of Leviticus was not always the chattel slavery most people think of from the 18th century. For fellow Israelites, it was typically a form of indentured servitude, often willingly entered into to pay off a debt. Just because something was reported to have happened in the Bible, doesn't always mean it condones it. I see you left off many of the newer passages about slavery that would refute your su…

[dead]
Post reply on HN