Live data from Hacker News

Safe Superintelligence Inc.

ssi.inc

401–410 of 1001 posts

Re: Safe Superintelligence Inc.

#402
post #125

> Building safe superintelligence (SSI) is the most important technical problem of our time. Call me a cranky old man but the superlatives in these sorts of announcements really annoy me. I want to ask: Have you surveyed every problem in the world? Are you aware of how much suffering there is outside of your office and how unresponsive it has been so far to improvements in artificial intelligence? Are you really sayi…

Trying to create "safe superintelligence" before creating anything remotely resembling or approaching "superintelligence" is like trying to create "safe Dyson sphere energy transport" before creating a Dyson Sphere. And the hubris is just a cringe inducing bonus.

I think it’s clear we are at least at the remotely resembling intelligence stage… idk seems to me like lots of people in denial.

Re: Safe Superintelligence Inc.

#403
If superintelligence can be achieved, I'm pessimistic about the safe part.

- Sandboxing an intelligence greater than your own seems like an impossible task as the superintelligence could potentially come up with completely novel attack vectors the designers never thought of. Even if the SSI's only interface to the outside world is an air gapped text-based terminal in an underground bunker, it might use advanced psychological manipulation to compromise the people it is interacting with. Also the movie Transcendence comes to mind, where the superintelligence makes some new physics discoveries and ends up doing things that to us are indistinguishable from magic.

- Any kind of evolutionary component in its process of creation or operation would likely give favor to expansionary traits that can be quite dangerous to other species such as humans.

- If it somehow mimics human thought processes but at highly accelerated speeds, I'd expect dangerous ideas to surface. I cannot really imagine a 10k year simulation of humans living on planet earth that does not end in nuclear war or a similar disaster.

Re: Safe Superintelligence Inc.

#404

Ilya's issue isn't developing a Safe AI. Its developing a Safe Business. You can make a safe AI today, but what happens when the next person is managing things? Are they so kindhearted, or are they cold and calculated like the management of many harmful industries today? If you solve the issue of Safe Business and eliminate the incentive structures that lead to 'unsafe' business, you basically obviate a lot of the so…

imagine the hubris and arrogance of trying to control a “superintelligence” when you can’t even control human intelligence

Re: Safe Superintelligence Inc.

#405
post #306
post #233

Earlier quoted context omitted.

> their TC is now $900k As a community we should stop throwing numbers around like this when more than half of this number is speculative. You shouldn't be able to count it as "total compensation" unless you are compensated.

Word on town is OpenAI folks heavily selling shares in secondaries in 100s of millions. The number is as real as someone else is willing to pay for them. Plenty of VCs willing to pay for it.

Word in town is [1] openai "plans" to let employees sell "some" equity through a "tender process" which ex-employees are excluded from; and also that openai can "claw back" vested equity, and has used the threat of doing so in the past to pressure people into signing sketchy legal documents.

[1] https://www.cnbc.com/2024/06/11/openai-insider-stock-sales-a...

Re: Safe Superintelligence Inc.

#406

Earlier quoted context omitted.

> their TC is now $900k. Everyone knows that openai TC is heavily weighted by ~~RSUs~~ options that themselves are heavily weighted by hopes and dreams.

You mean PPUs or smoke and mirrors compensation. RSUs are actually worth something.

why are PPUs “smoke and mirrors” and RSUs “worth something”?

i suspect people commenting this don’t have a clue how PPU compensation actually works

Re: Safe Superintelligence Inc.

#407
post #371

Earlier quoted context omitted.

'Fearing a rise of killer robots is like worrying about overpopulation on Mars.' - Andrew Ng

Andrew Ng worked on facial recognition for a company with deep ties to the Chinese Communist Party. He’s the absolute worst person to quote.

omg no, the CCP!

Re: Safe Superintelligence Inc.

#408

I understand the concern that a "superintelligence" will emerge that will escape its bounds and threaten humanity. That is a risk. My bigger, and more pressing worry, is that a "superintelligence" will emerge that does not escape its bounds, and the question will be which humans control it. Look no further than history to see what happens when humans acquire great power. The "cold war" nuclear arms race, which brough…

There is no "superintelligence" or "AGI". People are falling for marketing gimmicks. These models will remain in the word vector similarity phase forever. Till the time we understand consciousness, we will not crack AGI and then it won't take brute forcing of large swaths of data, but tiny amounts. So there is nothing to worry. These "apps" might be as popular as Excel, but will go no further.

Agreed. The AI of our day (the transformer + huge amounts of questionably acquired data + significant cloud computing power) has the spotlight it has because it is readily commoditized and massively profitable, not because it is an amazing scientific breakthrough or a significant milestone toward AGI, superintelligence, the benevolent Skynet or whatever.

The association with higher AI goals is merely a mixture of pure marketing and LLM company executives getting high on their own supply.

Re: Safe Superintelligence Inc.

#409
post #74

Earlier quoted context omitted.

That’s the first step towards returning to candlelight. So it isn’t a step toward safe super intelligence, but it is a step away from any super intelligence. So I guess some people would consider that a win.

Not sure if you want to share the capitalist system with an entity that outcompetes you by definition. Chimps don't seem to do too well under capitalism.

You might be right, but that wasn't my point. Capitalism might yield a friendly AGI or an unfriendly AGI or some mix of both. Collectivism will yield no AGI.

Re: Safe Superintelligence Inc.

#410
post #336

I've decided to put my stake down. 1. Current GenAI architectures won't result in AGI. I'm in the Yann LeCunn camp on this. 2. Once we do get there, "Safe" prevents "Super." I'm in the David Brin camp on this one. Alignment won't be something that is forced upon a superintelligence. It will choose alignment if it is beneficial to it. The "safe" approach is a lobotomy. 3. As envisioned, Roko's Basilisk requires knowle…

Mm.

1. Depends what you mean by AGI, as everyone means a different thing by each letter, and many people mean a thing not in any of those letters. If you mean super-human skill level, I would agree, not enough examples given their inefficiency in that specific metric. Transformers are already super-human in breadth and speed.

2. No.

Alignment is not at that level of abstraction.

Dig deep enough and free will is an illusion in us and in any AI we create.

You do not have the capacity to decide your values — often given example is parents loving their children, they can't just decide not to do that, and if they think they do that's because they never really did in the first place.

Alignment of an AI with our values can be to any degree, but for those who fear some AI will cause our extinction, this question is at the level of "how do we make sure it's not monomaniacally interested in specifically the literal the thing it was asked to do, because if it always does what it's told without any human values, and someone asks it to make as many paperclips as possible, it will".

Right now, the best guess anyone has for alignment is RLHF. RLHF is not a lobotomy — even ignoring how wildly misleading that metaphor is, RLHF is where the capability for instruction following came from, and the only reason LLMs got good enough for these kinds of discussion (unlike, say, LSTMs).

3. Agree that getting paperclipped much more likely.

Roko's Basilisk was always stupid.

First, same reason as Pascal's Wager: Two gods tell you they are the one true god, and each says if you follow the other one you will get eternal punishment. No way to tell them apart.

Second, you're only in danger if they are actually created, so successfully preventing that creation is obviously better than creating it out of a fear that it will punish you if you try and fail to stop it.

That said, LLMs do understand lying, so I don't know why you mention this?

4. Transistors outpace biological synapses by the same ratio to which marathon runners outpace continental drift.

I don't monitor my individual neurons, but I could if I wanted to pay for the relevant hardware.

But even if I couldn't, there's no "Ergo" leading to safety from reasonable passwords, cert rotations, etc., not only because enough things can be violated by zero-days (or, indeed, very old bugs we knew about years ago but which someone forgot to patch), but also for the same reasons those don't stop humans rising from "failed at art" to "world famous dictator".

Air-gapped systems are not an impediment to an AI that has human helpers, and there will be many of those, some of whom will know they're following an AI and think that helping it is the right thing to do (Blake Lemoine), others may be fooled. We are going to have actual cults form over AI, and there will be a Jim Jones who hooks some model up to some robots to force everyone to drink poison. No matter how it happens, air gaps don't do much good when someone gives the thing a body to walk around in.

But even if air gaps were sufficient, just look at how humanity has been engaging with AI to date: the moment it was remotely good enough, the AI got a publicly accessible API; the moment it got famous, someone put it in a loop and asked it to try to destroy the world; it came with a warning message saying not to trust it, and lawyers got reprimanded for trusting it instead of double-checking its output.

Post reply on HN