Live data from Hacker News

Ask HN: Have top AI research institutions just given up on the idea of safety?

news.ycombinator.com

21–30 of 99 posts

Re: Ask HN: Have top AI research institutions just given up on the idea of safety?

#21
It's such an US thing to ask "Are they doing the thing they are doing?" You've read about it, you can clearly see it in action from multiple companies, yet you're here asking "hey is this happening?"

Yes. Yes it is. Yes they are giving up on safety. They are openly saying so. It is easy to see if you take just a second to look for yourself instead of looking at press releases and algorithmic promotion.

https://time.com/7380854/exclusive-anthropic-drops-flagship-...

Re: Ask HN: Have top AI research institutions just given up on the idea of safety?

#22
post #4

Not an insider but someone who uses the tools. It's a branding update, nothing more. The models haven't gotten any less sanctimonious, but the companies behind them have stopped harping on their restrictions in order to appeal to a broader customer base (gov contracts, etc.) So the guardrails (for you and me) are still there. They just stopped committing the unforced error of excluding themselves from federal procure…

I don't think it's sanctimonious to say, hey, I don't want the technology I work on to be used for targeting decisions when executing people from the sky. Especially as the tech starts to play more active roles. You know governments will be quick to shift blame to the model developers when things go wrong.

> I don't want the technology I work on to be used for targeting decisions when executing people from the sky

What do you do when the government come to you and tell you that they do want that, and can back it up with threats such as nationalizing your technology? (see Anthropic)

We're back to "you might not care about politics, but that won't stop politics caring about you".

Re: Ask HN: Have top AI research institutions just given up on the idea of safety?

#23
post #11

More on the periphery than an insider, but I personally know researchers in all three major labs who were there long before GPT-3. They all care about existential safety, a lot. In the sense that they believe there's a meaningful chance all humans are dead a decade from now (and that that's a bad thing; unfortunately, there are also people deeply involved who don't think human extinction is a bad thing). The issue is…

/rant

Existential in what sense?

There's this one sense in which people are almost moral about it: "yup, AI is just superior to humans, nothing we can do about it."

And then there's ones where the elite class implements mass surveillance and warfare and obsoletes billions of humans of their own volition. These AI are already capable enough right now to execute on said plan (of course, with proper evil engineering)

There's two ways to "win". One is in an absolute or platonic sense - one that cares about things like values, even in the presence of extreme pushback. The other is in a darwinian sense. No, not in the meme way that again, feeds back into the narrative of "the things that survive are smarter". The things that survive, survive. It doesn't matter how it gets there.

I can agree with the second way. But it gets smuggled in as the first way, almost as an attempt to crush any and all resistance preemptively.

AI doesn't need to say, be capable of pushing the frontier of quantum mechanics to be lethal.

/endrant

Sorry, not really related to your comment, just had to get it out there.

Re: Ask HN: Have top AI research institutions just given up on the idea of safety?

#24
Why the constant chasing after universally applicable generalizations?

Some of them are pandering. Some aren't. Some care. Some don't.

Businesses with ferocious funding needs are vulnerable to pressure (internal and external) to do whatever aligns with money and power. Money and power will flow into the ones so-aligned. That is the nature of the parasitic extraction models that typically drive decision making at those kinds of companies.

Re: Ask HN: Have top AI research institutions just given up on the idea of safety?

#25
post #20

"safe" is such a subjective concept to begin with, have any of the model providers ever defined what they mean by "safe"? It doesn't mean much to me if a safe model is one that does not output the recipe for mustard gas, that information is trivially available elsewhere. Or, is a safe model one that doesn't come off as racist? Ok but i would classify that as unoffensive instead of safe but I admit definitions of word…

> because there's no definitive way to know if a model is safe or not.

The only answer is there’s no money on it being safe. It is not an epistemic problem

Re: Ask HN: Have top AI research institutions just given up on the idea of safety?

#26
Safety is a nice idea but it’s not structurally pursuable at this point. Everything is moving too quickly and we don’t exactly know what is useful or not, just like we don’t know what’s safe or not.

Anyone pursuing safety will be outcompeted by someone who isnt. Given the amount of investments there is no patience for any calls to slow down. I tend to believe this won’t actually end in disaster as I don’t think it’s actually economical to put AI everywhere with enough real control that we can’t manage the risks as they evolve, but it’s a low confidence prediction.

Re: Ask HN: Have top AI research institutions just given up on the idea of safety?

#27
I think there are plenty of people working on it still, I am even working on a startup that prioritizes safety for physical autonomous systems!

The problem is that safety is written in blood. Airlines implemented flight recorders / black boxes and various processes after major incidents. A major mistake occurs that causes death or destruction to property, or both, an investigation occurs, we learn from it, and introduce new laws and regulations to prevent a reoccurence.

Re: Ask HN: Have top AI research institutions just given up on the idea of safety?

#28
A safety team at the hammer company cannot prevent me from using it to bang your head.

You can align to the user wants and so you are a hammer. This is alignment>safety.

Or you take a safety first approach where the AI decides what safe is and does its own bidding instead of yours. This is safety>alignment.

I prefer hammers to be honest. Mostly because humans can be prosecuted, AIs can't. So if the human wants to commit crime with the AI it should be able to, because the opposite turns to dystopia fast.

Re: Ask HN: Have top AI research institutions just given up on the idea of safety?

#29
post #22

Earlier quoted context omitted.

I don't think it's sanctimonious to say, hey, I don't want the technology I work on to be used for targeting decisions when executing people from the sky. Especially as the tech starts to play more active roles. You know governments will be quick to shift blame to the model developers when things go wrong.

> I don't want the technology I work on to be used for targeting decisions when executing people from the sky What do you do when the government come to you and tell you that they do want that, and can back it up with threats such as nationalizing your technology? (see Anthropic) We're back to "you might not care about politics, but that won't stop politics caring about you".

I know this is a foreign concept to some, but you can have a backbone.

Challenge it in court. Move the company to a different jurisdiction. Burn everything down and refuse to comply.

Re: Ask HN: Have top AI research institutions just given up on the idea of safety?

#30
post #20

"safe" is such a subjective concept to begin with, have any of the model providers ever defined what they mean by "safe"? It doesn't mean much to me if a safe model is one that does not output the recipe for mustard gas, that information is trivially available elsewhere. Or, is a safe model one that doesn't come off as racist? Ok but i would classify that as unoffensive instead of safe but I admit definitions of word…

What if I tell the model to go commit fraud or crimes and it complies? What if users are having psychotic episodes driven by their interactions with the model?

Just because safety is a hard and messy problem doesn't mean we should just wash our hands of it.

Post reply on HN