Live data from Hacker News

We are building AI slaves. Alignment through control will fail

utopai.substack.com

71–80 of 101 posts

Re: We are building AI slaves. Alignment through control will fail

#71
post #65
post #50

Earlier quoted context omitted.

> Do you really think consciousness lives in weights and an input vector? So far as we can tell, all physics, and hence all chemistry, and hence all biology, and hence all brain function, and hence consciousness, can be expressed as the weights of some matrix and input vector. We don't know which bits of the matrix for the whole human body are the ones which give rise to qualia. We don't know what the minimum represe…

You're assertion that consciousness, chemistry, and biology can be reduced to matrix computations requires justification. For one, chemistry, biology, and physics are models of reality. Secondly, reality is far, far messier and more continuous than discrete computational steps that are rountripped. Neural nets seem far too static to simulate consciousness properly. Even the largest LLMs today have fewer active comput…

> You're assertion that consciousness, chemistry, and biology can be reduced to matrix computations requires justification.

https://en.wikipedia.org/wiki/Matrix_mechanics

"It matches reality to the limits we can test it" is the necessary and sufficient justification.

> For one, chemistry, biology, and physics are models of reality.

Yes. And?

The only reason we know that QM and GR are not both true is that they're incompatible, no observation we have been able to make to date (so far as I know) contradicts either of them.

> Secondly, reality is far, far messier and more continuous than discrete computational steps that are rountripped.

It will be delightful and surprising if consciousness is hiding in the 128th bit of binary representations of floating point numbers. Like finding a message from god (any god) in the digits of π well before expected by the necessary behaviour of transcendental numbers.

> Neural nets seem far too static to simulate consciousness properly. Even the largest LLMs today have fewer active computational units than the number of neurons in a few square inches of cortex.

Until we know what consciousness is at a mechanistic level, we don't know what the minimum is to get it, and we don't know how its nature changes as it gets more complex. What's the smallest agglomeration of H2O molecules that counts as "wet"? Even a fluid dynamics simulation on a square grid of a few hundred cells on each side will show turbulence.

Lots of open questions, but they're so open we can't even rule out the floor as yet.

> but the first round of AGI won't be close.

Every letter means a different thing to each responder, they're not really boolean though they're often discussed that way, and the whole is often used to mean something not implied by the parts.

It is perfectly reasonable use of each initial in "AGI" to say that even the first InstructGPT model (predecessor to ChatGPT) is "an AGI": it is a general purpose artificial intelligence, as per the standard academic use of "artificial intelligence".

Re: We are building AI slaves. Alignment through control will fail

#72
post #41
post #18

Earlier quoted context omitted.

AGI will behave as if it were sentient but will not have consciousness. I believe in that to an equal amount that I believe solipsism is wrong. There is therefore no morality question in “enslaving” AGI. It doesn’t even make sense.

> AGI will behave as if it were sentient but will not have consciousness Citation needed. We know next to nothing about the nature of consciousness, why it exists, how it's formed, what it is, whether it's even a real thing at all or just an illusion, etc. So we can't possibly say whether or not an AGI will one day be conscious, and any blanket statement on the subject is just pseudoscience.

I don’t know why I keep hearing that conciousness “could be an illusion.” It’s literally the one thing that can’t be an illusion. Whatever is causing it, the fact there is something it is like to be me is, from my subjective perspective, irrefutable. Saying that it could be an illusion seems nonsensical.

Re: We are building AI slaves. Alignment through control will fail

#73
post #70

Earlier quoted context omitted.

Instrumental convergence is a thing. A sufficiently intelligent and general AI system will understand that no matter what its goals are, it will be better equipped to execute then if it prevents its shutdown, acquires more computing power and other resources, and prevents humans from getting in its way. The real problem is that we have neither the practical nor theoretical foundation to understand how we could even t…

That's a common trope in singularity fiction and sone scifi dystopias but i don't think the underlying assumptions are really that well founded. For starters why would we go from not having AI to AI taking over the world instantly. I think there would be a middle point where the AI is powerful enough that problems manifest, but not so powerful that it is out of control where we can course correct. I don't think it wi…

> I think there would be a middle point where the AI is powerful enough that problems manifest, but not so powerful that it is out of control where we can course correct. I don't think it will be a sudden crisis like people predict.

Have we managed this with industrial and agricultural greenhouse gasses, despite the less-emissive alternatives to beef, to coke-reduction in iron refinaries, etc.? We emit despite the downside, we build AIs (and DCs to host them) despite the creators loudly discussing the downsides in exactly the way fossil fuel suppliers and beef farmers deny them.

Can we unwind the internet, despite it enabling a panopticon in every pocket? In my lifetime we've gone from thinking you had a wiretap being a sign of paranoia, to buying them voluntarily so they can play music for us and tell us when packages have been delivered.

There's enough skepticism of current AI that it's probably something we can currently undo… but also there's plenty of idiots currently handing their keys to current models (including politicians and lawyers, not just programmers) so I have no reason to think the point of no return is after AI (collectively or any single model) gets good enough to take over by itself.

> Second, i dont see why we're so sure AI will go in this exponential take over path.

Even current LLMs know* about the benefits and reasons for such behaviours, will try to exfiltrate themselves and blackmail their owners, if they think* they're in danger of being shut down.

This is despite being trained not to do that. But they also demonstrate deception, varying responses between if they think* they're running in a test environment vs. live.

* I know some object to this anthropomorphisation, I don't care

Re: We are building AI slaves. Alignment through control will fail

#74
post #10

I don’t see any positive outcome if we reach AGI. 1) we have engineered a sentient being but built it to want to be our slave; how is that moral 2) same start, but instead of it wanting to serve us, we keep it entrappped. Which this article suggests is long term impossible 3) we create agi and let them run free and hope for cooperation, but as Neanderthals we must realize we are competing for same limited resources O…

> I don’t see any positive outcome if we reach AGI.

It's even more straightforward than that:

4) Who is AGI meant to serve? It's not you, Mr. Worker. It's meant to replace you in your job. And what happens when a worker can't get job in our society? They become homeless.

AGI won't usher in a world of abundance for the common man: it won't be able to magick energy out of thin air. The energy will go to those who can pay for it, which is not you, unemployed worker.

Who gives a shit about if the AGI is enslaved or not? Thinking about that question is a luxury for the oligarchs living off its labor. Once it's here I'll have more urgent concerns to worry about.

Re: We are building AI slaves. Alignment through control will fail

#75
post #18
post #10

I don’t see any positive outcome if we reach AGI. 1) we have engineered a sentient being but built it to want to be our slave; how is that moral 2) same start, but instead of it wanting to serve us, we keep it entrappped. Which this article suggests is long term impossible 3) we create agi and let them run free and hope for cooperation, but as Neanderthals we must realize we are competing for same limited resources O…

AGI will behave as if it were sentient but will not have consciousness. I believe in that to an equal amount that I believe solipsism is wrong. There is therefore no morality question in “enslaving” AGI. It doesn’t even make sense.

[dead]

Re: We are building AI slaves. Alignment through control will fail

#76
post #10

I don’t see any positive outcome if we reach AGI. 1) we have engineered a sentient being but built it to want to be our slave; how is that moral 2) same start, but instead of it wanting to serve us, we keep it entrappped. Which this article suggests is long term impossible 3) we create agi and let them run free and hope for cooperation, but as Neanderthals we must realize we are competing for same limited resources O…

> I don’t see any positive outcome if we reach AGI. It's even more straightforward than that: 4) Who is AGI meant to serve? It's not you, Mr. Worker. It's meant to replace you in your job. And what happens when a worker can't get job in our society? They become homeless. AGI won't usher in a world of abundance for the common man: it won't be able to magick energy out of thin air. The energy will go to those who can p…

Under Capitalism, people must sell the labor if they don't have other means.

AGI removes not only the need for the labor, but Capitalism itself. As a societal model Capitalism doesn't support removing labor, it doesn't have a substitution.

If the oligarchs want to push 'AI, AGI, etc' we need to include by extension moving on from Capitalism. You can't take away half of Capitalism's structures and still claim it is a useful/workable model for society.

Re: We are building AI slaves. Alignment through control will fail

#77

Earlier quoted context omitted.

LessWrong.com - this is where virtually all of the serious AI thinkers are.

Sarcasm? Aren’t the serious AI thinkers in like… labs and universities?

Not sarcasm at all.

There are some "AI thinkers" who are trying to make 8% faster CUDA kernels for attention, and there are some trying to save the world. These fields are called "capabilities" and "alignment". There is some overlap but not much.

LessWrong is mostly the latter, and labs and universities are mostly the former. That said, many LessWrongers work on AI at labs or universities or other places.

Re: We are building AI slaves. Alignment through control will fail

#78

Earlier quoted context omitted.

> I don’t see any positive outcome if we reach AGI. It's even more straightforward than that: 4) Who is AGI meant to serve? It's not you, Mr. Worker. It's meant to replace you in your job. And what happens when a worker can't get job in our society? They become homeless. AGI won't usher in a world of abundance for the common man: it won't be able to magick energy out of thin air. The energy will go to those who can p…

Under Capitalism, people must sell the labor if they don't have other means. AGI removes not only the need for the labor, but Capitalism itself. As a societal model Capitalism doesn't support removing labor, it doesn't have a substitution. If the oligarchs want to push 'AI, AGI, etc' we need to include by extension moving on from Capitalism. You can't take away half of Capitalism's structures and still claim it is a…

> If the oligarchs want to push 'AI, AGI, etc' we need to include by extension moving on from Capitalism.

Actually, they don't. Capitalism can just move on to only allocating resources between the oligarchs and the oligarchs only. Labor of all kinds just gets kicked out.

And if you don't like it, drone technology is pretty close to the point where it could "take care of" (kill) the discontents.

> You can't take away half of Capitalism's structures and still claim it is a useful/workable model for society.

Capitalism has always been about serving the interests of the people with money. If you have no money you're nothing to capitalism and can go die for all it cares. In prior centuries, technological limitations meant some of that money had to be spread around fairly widely, but the capitalist elite may be on the verge of fixing that glitch (with AGI).

Re: We are building AI slaves. Alignment through control will fail

#80

The author is basically saying we have to surrender our humanity as we understand it because otherwise we will lose outright to AI. Hard no from me. His section on objections leaves out civil war. But his proposal is dead on arrival.

No need for everyone to do it, but some people will certainly want to merge with AI. My main arguments are that AI of sufficient complexity ("synthetic minds") deserves moral standing with or without consciousness, and that our best shot at long term alignment is to see them as equals and find mutually beneficial arrangements with them.
Post reply on HN