Live data from Hacker News

Anthropic's Claude is said to improve on ChatGPT, but still has limitations

techcrunch.com

1–10 of 58 posts

Re: Anthropic's Claude is said to improve on ChatGPT, but still has limitations

#2
> Anthropic started with a list of around ten principles that, taken together, formed a sort of “constitution” (hence the name “constitutional AI”). The principles haven’t been made public, but Anthropic says they’re grounded in the concepts of beneficence (maximizing positive impact), nonmaleficence (avoiding giving harmful advice) and autonomy (respecting freedom of choice).

This is giving me very strong Asimov's "three laws of robotics" vibes:

First Law

A robot may not injure a human being or, through inaction, allow a human being to come to harm.

Second Law

A robot must obey the orders given it by human beings except where such orders would conflict with the First Law.

Third Law

A robot must protect its own existence as long as such protection does not conflict with the First or Second Law.

Re: Anthropic's Claude is said to improve on ChatGPT, but still has limitations

#4

> Anthropic started with a list of around ten principles that, taken together, formed a sort of “constitution” (hence the name “constitutional AI”). The principles haven’t been made public, but Anthropic says they’re grounded in the concepts of beneficence (maximizing positive impact), nonmaleficence (avoiding giving harmful advice) and autonomy (respecting freedom of choice). This is giving me very strong Asimov's "…

> or, through inaction, allow a human being to come to harm.

This has always felt like a gaping hole to me. It seems like to work it would have to a) always make perfect predictions of the future, and b) agree with relevant humans what "harm" is.

Re: Anthropic's Claude is said to improve on ChatGPT, but still has limitations

#5

> Anthropic started with a list of around ten principles that, taken together, formed a sort of “constitution” (hence the name “constitutional AI”). The principles haven’t been made public, but Anthropic says they’re grounded in the concepts of beneficence (maximizing positive impact), nonmaleficence (avoiding giving harmful advice) and autonomy (respecting freedom of choice). This is giving me very strong Asimov's "…

> First Law

> A robot may not injure a human being or, through inaction, allow a human being to come to harm.

Hmmm I wonder about trolley problems

Re: Anthropic's Claude is said to improve on ChatGPT, but still has limitations

#6
In a strange twist of events... their massive Series B round was led by SBF

"The [$580M] Series B follows the company raising $124 million in a Series A round in 2021. The Series B round was led by Sam Bankman-Fried, CEO of FTX. The round also included participation from Caroline Ellison, Jim McClave, Nishad Singh, Jaan Tallinn, and the Center for Emerging Risk Research (CERR)."

https://www.anthropic.com/news/announcement

Re: Anthropic's Claude is said to improve on ChatGPT, but still has limitations

#7

> Anthropic started with a list of around ten principles that, taken together, formed a sort of “constitution” (hence the name “constitutional AI”). The principles haven’t been made public, but Anthropic says they’re grounded in the concepts of beneficence (maximizing positive impact), nonmaleficence (avoiding giving harmful advice) and autonomy (respecting freedom of choice). This is giving me very strong Asimov's "…

> First Law > A robot may not injure a human being or, through inaction, allow a human being to come to harm. Hmmm I wonder about trolley problems

You may already know, but _most_ of Asimov's stories with the three laws are heavily based on problems with said laws, and clever loopholes of various sorts.

So it's always hilarious when people use them unironically/uncritically in other contexts.

Re: Anthropic's Claude is said to improve on ChatGPT, but still has limitations

#8
post #4

> Anthropic started with a list of around ten principles that, taken together, formed a sort of “constitution” (hence the name “constitutional AI”). The principles haven’t been made public, but Anthropic says they’re grounded in the concepts of beneficence (maximizing positive impact), nonmaleficence (avoiding giving harmful advice) and autonomy (respecting freedom of choice). This is giving me very strong Asimov's "…

> or, through inaction, allow a human being to come to harm. This has always felt like a gaping hole to me. It seems like to work it would have to a) always make perfect predictions of the future, and b) agree with relevant humans what "harm" is.

That's sort of the point of his stories. The laws do not work, and there is no set of laws that could work perfectly. Any set of rules as simple as this applied to something as complex as humanity will always have a mountain of loopholes.

Re: Anthropic's Claude is said to improve on ChatGPT, but still has limitations

#9

> Anthropic started with a list of around ten principles that, taken together, formed a sort of “constitution” (hence the name “constitutional AI”). The principles haven’t been made public, but Anthropic says they’re grounded in the concepts of beneficence (maximizing positive impact), nonmaleficence (avoiding giving harmful advice) and autonomy (respecting freedom of choice). This is giving me very strong Asimov's "…

Maybe the ML model would start a market maker and exchange in the Bahamas too.

Re: Anthropic's Claude is said to improve on ChatGPT, but still has limitations

#10

In a strange twist of events... their massive Series B round was led by SBF "The [$580M] Series B follows the company raising $124 million in a Series A round in 2021. The Series B round was led by Sam Bankman-Fried, CEO of FTX. The round also included participation from Caroline Ellison, Jim McClave, Nishad Singh, Jaan Tallinn, and the Center for Emerging Risk Research (CERR)." https://www.anthropic.com/news/announc…

When someone pays you with stolen funds, aren't you (morally, if not legally) obliged to pay it back to the victims?
Post reply on HN