Live data from Hacker News

Meet “Claude”: Anthropic’s rival to ChatGPT

scale.com

151–160 of 165 posts

Re: Meet “Claude”: Anthropic’s rival to ChatGPT

#151

Earlier quoted context omitted.

If someone gives a language model the capability for unfettered interaction with the physical world, and they are not liable for the consequences, then no safety feature of Claude can save us. And if they are liable, that is the primary mechanism which will ensure they take necessary steps to avoid negative consequences.

How many years until we have an AI decision making system that requires human approval convincing the human its decision is best, ending in catastrophe? (Sorry, posed merely for thought and not for dismissal of current/future achievements.) Makes me also wonder if you had two polarized bots arguing/discussing with each other, how long would it take for one to convince the other?

To me it really comes down to, can it have legal liability or not. The idea of AI convincing humans and vice versa is a bit abstract. To some extent I know I pretend I have free will, and I know I pretend the AI is meaningfully different from me in that it can't really have intent, it just generates output from input. But I'm pretty sure that I'm fundamentally the same as the AI in that sense.

However, the AI is but one of many agents trying to convince me, and there are many other things also in the mix. It is a bit like banning the knowledge of the second world war in fear that someone will learn that fascism is possible.

For this whole song and dance we call civilization as we know it to function, we have to say humans are accountable for their actions, with some well defined exceptions like duress and insanity. Save those exceptions, I can't absolve my liability by saying something, or someone, convinced me to do something by talking to me.

If ChatGPT comes to me with a gun to my head and tells me to do something, that is a different matter, but then liability shifts to whoever gave ChatGPT a gun, even if ChatGPT convinced that person to give it a gun by making a really good written argument.

Re: Meet “Claude”: Anthropic’s rival to ChatGPT

#152
Is this not a superficial attempt at saftey?

I would like my AI system to tell me how to hotwire a car if I am curious about how that works.

I would like my AI system to give me a detailed step by step car hotwire walkthrough if I am in a physically abusive relationship and my kids and I only have 30 minutes to try to hotwire the car and escape a remote area for safety.

I do not want AI systems to create children's books in the style of authors that I know, for the purposes of selling books and reducing my friends' ability to have a happy productive life. Especially because it was trained on their work. I want my friends to be happy, and I have had some friends commit suicide. So maybe improving human happiness is a saftey concern, and generating kids books is not safe. But that doesn't look like "safety" from a superficial point of view.

The only way for an AI to be able to make judgements on safety is for it to have general intelligence and some life experience (like we do). Because it needs to figure out context to know if it should be telling a particular person how to hotwire a car.

I am being very dismissive because I don't see this as being a perfect solution, and it is easy to see why. But maybe someone who works on this can explain how an imperfect solution still has value? I am open to that possibility.

Maybe self-reflection and self-tuning is of general value - even if it only superficially addresses safety concerns in a 1 dimensional way.

Perhaps these techniques can be used on something other than safety.

Re: Meet “Claude”: Anthropic’s rival to ChatGPT

#153

Hello HN — I’m the coauthor of this post. You may remember me as that guy who spent most of 2022 posting GPT-3 screenshots to Twitter, most famously prompt injection and “You are GPT-3”. Happy to answer any questions about Claude that I can.

[I mean no bad faith in this comment, I'm a fan of yours.] Why answer questions about harmlessness/safety in such a roundabout way? Both OpenAI and Anthropic are clear about what words like "safe" are intended to mean: a stepping stone to "AI does not kill all people when given control". Avoiding to state this clearly only invites unnecessary culture war disagreements in every discussion about these models.

Maybe you’re right. It’s partially laziness on my part — it takes a while to explain long-term issues, and those who are inclined to care about them are generally aware of who started Anthropic and why.

Re: Meet “Claude”: Anthropic’s rival to ChatGPT

#154
post #38

Earlier quoted context omitted.

I have to look that Geordi episode up. But I am mostly familiar with the later TNG era star trek, so I didn’t know it was written as self-aware in the early days. Some episodes do feature “bugs” where holographic actors become aware being in a program/being an actor. The episode where an Irish town program has run too long on Voyager comes to mind. (Edit: I do wonder if the holographic actors are somehow sandboxed co…

The episode is "The Mind's Eye". But I think the Star Trek writers just didn't understand computers very well. In the future, making duplicate copies of data seems to be impossible. When you copy a file from one device to another, or one ship to another, or transmit it to a planet, it seems to disappear from the source. This is a particularly common weirdness in Voyager, where duplicating holographic programs is appa…

Thanks for sharing the name of the episode. Something to watch today.

Maybe they figured out blockchain unique data structures that can’t be copied ;)

Re: Meet “Claude”: Anthropic’s rival to ChatGPT

#155

Earlier quoted context omitted.

For humans yes, I am not saying one-shot learning would be possible for undocumented indigenous languages but few shot language acquisition in cases of a single surviving speaker is something that I would consider highly probable . This hypothesis relies heavily on the nature of variational learning in latent space and observations about human languages. It is of course possible that some ethnicity would have a langu…

In encryption it's generally impossible to decrypt a 1 to many hash. You can do some clever things (correlating and combining other data) but if you're just looking at some hash that could be an infinite number of other things, you're just out of luck. I'll take the extreme position that language translation is an unsolvable problem because of this exact phenomena. There was a recent case where a politician was accus…

I don’t think the pigeonhole principle applies to human languages though

Re: Meet “Claude”: Anthropic’s rival to ChatGPT

#156
post #91

Earlier quoted context omitted.

You're welcome to try GPT3 before instruction tuning. It doesn't work at all. "Uncensored" AIs especially don't work for women because they'll immediately start writing erotica.

I don't have an issue with instruction tuning, but it's disengenious to pretend the biases inherent in the instruction tuning are a good thing and 'ethics'.

It's always got biases (it's made of them) so some biases must be better than others.

https://spetharrific.tumblr.com/post/26600309788/sussman-att...

The untuned model isn't "an average of all opinions on Earth" or anything either.

Re: Meet “Claude”: Anthropic’s rival to ChatGPT

#157
post #96

> Claude feels not only safer but more fun than ChatGPT. I may be the minority here, but I really don't concern myself with ChatGPT safety and I am not entirely sure what the reason is why people are very worried about its safetly. It is safer than most things I have in my house, including a kettle, a saw, a hammer, a screwdriver, my actual PC, every kitchen appliance I have. Of course it can be misused, like any too…

I hope someone makes one of these things that has been trained without any concern for 'safety' or propriety, just for the sake of comparison.

I would note that

1. ChatGPT is just OpenAI's default model with a specific prompt and interaction-buffer rewrite model; and

2. you can sign up, for free, for access to OpenAI's API (https://openai.com/api/); which, among other things, gives you full (free-daily-API-credit-quota limited) access to an API "playground" frontend offering interaction with the exact same model ChatGPT is powered by — plus other models as well — all without any fixed prompt or forced interaction UX. (In web-dev terms, if ChatGPT is like a REST API, this playground is like a SQL fiddle for the DB that the REST API is backed by.)

(Why does this exist? Because the point of OpenAI's API playground is to test prompts and interaction models for the AI apps you're building yourself on top of their API; and you couldn't very well build apps with your own prompts and interaction models, if OpenAI was already imposing a prompt and interaction model upon you.)

Re: Meet “Claude”: Anthropic’s rival to ChatGPT

#158

Earlier quoted context omitted.

> the person who will prospectively coerce the language model into being an asshole? The language models are trained on internet content already. If you ignore ethics, it means you're feeding it bias and racism along with everything else. Chatgpt is not blatant about it normally but I'm sure you've seen examples like "write a function that takes a race argument and returns length of prison sentence" where it obviousl…

I'm sure by now there is enough content on the internet that says racism is bad. It's got to outnumber racist things said online by like a million to one. Why not just let the model learn what it learns? If people care so much about racism then that should reflect naturally in the model.

Models don't learn to reason about things. They learn to rehash things in new context. If you train the model on both racist and anti-racist content, you'll get it repeating both racist and anti-racist ideas.

Re: Meet “Claude”: Anthropic’s rival to ChatGPT

#160
post #32

Here is a fun example of what it can do: https://twitter.com/jayelmnop/status/1612243602633068549 .

The "This Title Is Now Longer Than The Actual Movie" gag feels a bit too much like something ripped from the training set for me. I'm willing to be amazed though.
Post reply on HN