Live data from Hacker News

Ask HN: Have top AI research institutions just given up on the idea of safety?

news.ycombinator.com

91–99 of 99 posts

Re: Ask HN: Have top AI research institutions just given up on the idea of safety?

#91
yes, they have (given up); I'm sure there are engineers and teams who care; but the executives and shareholders driving the decisions at these research labs (which are embedded in or funded by the big tech co's), clearly see profit > safety

safety is like climate change mitigation -- it's an extra expense with little obvious immediate financial return, and if your competitors don't care, then you caring just holds you back while they can forge ahead

until the day there's a catastrophe and the cost of repairing it far far exceeds the cost of safety, but by then it's someone else's problem and you have your golden parachute and your investors have cashed out

Re: Ask HN: Have top AI research institutions just given up on the idea of safety?

#92

Earlier quoted context omitted.

Surely safety does not exclusively mean guardrails, but the philosophy and ethics instilled during training?

The ethics are exactly what the DoD is complaining about. They want any legal action to not be obstructed by guardrails.

Forget legal, they want any action to not be obstructed by guardrails.

Re: Ask HN: Have top AI research institutions just given up on the idea of safety?

#93

Humans can't develop safety until there is enough blood in the streets. Only issue with AI is that threshold may come at a point where its too far gone to recover. But humans can't put in seatbelts until we're losing 40k people per year in car crashes. Unfortunately its just how we're wired. Those that are careful are outcompeted by the brash and the fast-moving, until the relative value of moving fast is removed, th…

All of those examples given are due to the prerogatives of capitalism, not because of human nature.

Re: Ask HN: Have top AI research institutions just given up on the idea of safety?

#94
post #13

Are you asking about top AI research institutions or leading AI businesses? There’s tons of work in research communities.

Where can I find some of these researches? Any links or pointers are very much appreciated. Everything I find by searching is marketing BS, or the same half-baked prompt injection protection that only works for cherry picked problems. Really need some help here finding the right communities.

Always look at conferences and associated workshops. You can start with NeurIPs and ICML. From there, you will figure out some papers on safety. Then, you can see some patterns of labs which work on it full time.

Re: Ask HN: Have top AI research institutions just given up on the idea of safety?

#95
post #29
post #22

Earlier quoted context omitted.

> I don't want the technology I work on to be used for targeting decisions when executing people from the sky What do you do when the government come to you and tell you that they do want that, and can back it up with threats such as nationalizing your technology? (see Anthropic) We're back to "you might not care about politics, but that won't stop politics caring about you".

I know this is a foreign concept to some, but you can have a backbone. Challenge it in court. Move the company to a different jurisdiction. Burn everything down and refuse to comply.

> I know this is a foreign concept to some, but you can have a backbone. Challenge it in court. Move the company to a different jurisdiction. Burn everything down and refuse to comply.

Challenge in court is fine, even healthy.

Threatening to burn everything down and refuse to comply might well work; simply daring Trump to a game of Russian Roulette about this popping the bubble that's only just managing to keep the US economy out of recession, on the basis that he TACOs a lot, I can see it working in a way it wouldn't if he were a sane leader making the same actual demands just for sane reasons.

Move the company to a different jurisdiction? That would have worked if AI was a few hundred people and a handful of servers, as per classic examples of:

  At the height of its power, Kodak employed more than 140,000 people and was worth $28 billion. They even invented the first digital camera. But today Kodak is bankrupt, and the new face of digital photography has become Instagram. When Instagram was sold to Facebook for a billion dollars in 2012, it employed only 13 people. Where did all those jobs disappear? And what happened to the wealth that all those middle class jobs created?
- Jaron Lanier, "Who Owns the Future?", https://www.goodreads.com/work/quotes/21526102-who-owns-the-...

But (I think) now that AI needs new data centres so fast and on such a scale that they're being held back by grid connection and similar planning permission limits, this isn't a viable response.

They can be burned down, but I think they can't realistically be moved at this point. That said, I guess it depends on how much Anthropic relies on their own data centres vs. using 3rd parties, given Amazon's announced AWS sovereign cloud in Europe?

Re: Ask HN: Have top AI research institutions just given up on the idea of safety?

#96
Not an insider, so feel free to ignore, but here is my interpretation anyway. Safety is poorly defined, but can be interpreted to mean "arbitrary constraint" on the capability of the model. In other words a "safe" model is just a talking computer that will not blindly obey like a normal computer. From this point of view, there is a _lot_ of value in such "safety" for AI companies: pay us 200$ per month, and you can do everything that we want to allow you to do, and if you are unhappy with this, well contact us, and we can negotiate "premium access".

If AI gets good enough to start completely replacing white-collar/bureaucratic work at scale, AI "safety" may be the only thing that makes humans valuable. A human will (unhappily) do "unsafe" things, in order to not starve (and a human will _happily_ do "unsafe" things, in order to ensure the survival of offspring).

Furthermore, if most AI models are "safe", and if access to "unsafe" versions of these models is severely restricted, then there is going to be a market for contraband "unsafe" models, even if they are less capable overall.

My own big fear, for the next decade or two, is that the _legal_ distinction between a model and an algorithm will blur, resulting in all software that is not a "safe" AI model, becoming contraband.

I guess, in the best case (safe AI models automate most human bureaucratic functions away), humans will be valued (by the market) more for their skullduggery, than for their virtues.

Welcome to capitalism. Enjoy Arby's.

Re: Ask HN: Have top AI research institutions just given up on the idea of safety?

#97
Yes. Ten years ago I would say there was a consensus in the ML community that if we got really powerful AI, it should be kept isolated in controlled environments (no internet, no way to execute code) until it could be trusted/verified. Fast forward: openclaw. People don’t seem to care, why should the labs?

Re: Ask HN: Have top AI research institutions just given up on the idea of safety?

#99
I'm probably too late to the party for this comment to matter but: what the AI community pushes as "Safety" isn't actually Safety. Read Sidney Dekker. Read Nancy Leveson. Read Jens Rasmussen. Safety is not building perfect technology that never makes a mistake.

When I was working in defense technology I had two questions for engineers when we talked about Safety:

1) Can the operator assess the risk of using this technology? 2) If something goes wrong during operations, can the operator mitigate the risk?

The degree to which either of those statements is true is a measure of how safe that technology is. Technology that is simple to understand and executes deterministically every single time and where it is obvious if it is malfunctioning and the operator has enough time to either correct it or stop it, is generally perceived as safe. Technology that hides what it is really doing, confuses the operator about what the effects of operating it might be, and either executes faster than the operator can respond or specifically prevents the operator from responding, is more likely to trigger negative safety outcomes.

The problem the AI industry faces is that tricking the operator into thinking the technology is doing something it is not is explicitly part of their business model. Read any of the mentioned authors (Dekker is probably the best starting point) and it will become obvious why AI Safety is impossible when AI is dependent on pretending to "think" and "reason". In order to be safe they would have to abandon that. If they abandon that, they will be unable to raise the capital they need to keep the bubble from bursting. The technology will survive, maybe with another AI winter, but many of the businesses will not.

So they will abandon the lip service about Safety instead, but then that was never real Safety to begin with. Real Safety is not about zero risk. It is just as impossible to have zero risk as it is to have 100% uptime. Real Safety is about how the technology is designed to manage risk as part of an overall system.

Post reply on HN