Live data from Hacker News

Who Aligns the Aligners?

prestonbyrne.com

81–90 of 104 posts

Re: Who Aligns the Aligners?

#81
The mental model here is that AI is a powerful tool whose main purpose is to outcompete others, increase power or provide a defense. This is not the right model!

The biggest risk is not that someone discovered a big new weapon, which by the way

> No one person, or one company, knows the answer and history is no guide

are you kidding me? It's a problem as old as time.

This is not nuclear power though. The bigger problem is a scenario where AI increasingly acts like bacteria, doing things autonomously, things no one really understands, things we think are confusing and weird. And then just by chance, some bacteria end up in some human lung cells and they turn out to be an excellent place to grow and multiply. Why is this worse? Because weird shit does not seem threatening, and it almost always isn't, so nobody tries to stop it.

Also, the argument in the article seems to be "obviously AI development is not a particularly urgent global threat," and therefore these downsides (risk of collusion, misuse of state power, etc) are the most critical factors? I guess, uh, yeah, if there's no benefit to taking these actions we shouldn't take them.

Re: Who Aligns the Aligners?

#82
post #2

> History does provide a great deal of guidance, however, about the use and misuse of government power. It tells us that the state is in fact likely the worst possible custodian for the most powerful publication and data analysis technologies. Like, no. The worst custodian is someone like Thiel or Musk or Altman. > In foreign countries, particularly the United Kingdom and Europe, where national governments have fewer…

And they also send people to jail for twitter posts. Ill take good ole USA where at least I know I am free, over the tyranny of a parliament. Look at how France and Germany are responding to democratically elected parties who are vaguely right wing, they pull out all of the stops of suppressing them.

Ah yeah, two popular right wing lies - AfD is "vaguelly right" and UK is horrible tyrrany.

Re: Who Aligns the Aligners?

#83
post #5

> Nobody knows the answer and history is no guide, save that apocalyptic predictions about new technologies have, to date, all been wrong. If any past apocalyptic prediction had been true, you would not have been around to write this sentence and we would not have been around to read it. It’s not a very convincing argument when we cannot in principle observe the counterfactual.

Apocalyptic predictions normally get everyone into gear to prevent the threat. Not a great argument for "let's see what happens if we just sit around this time"

Re: Who Aligns the Aligners?

#84
post #17
post #8

> Eliezer Yudkowsky confidently asserting today that if we do not institute immediate global techno-communism, instituting draconian government control over speech and publication of a type never before seen in any Western society, we are all going to die Nice straw man... > There is no evidence that this will happen. No, non at all... Presumably if there was a risk we'd be seeing the majority of researchers at AI la…

`p(doom)` is ridiculous nonsense. How do people even come up with those numbers? Some people just say "I think it's 20% that AI will wipe us out" but how did they arrive at this number? Yeah they just made it up out of thin air, that's right. Complete fabrication without any evidence to back up such a probability. It's not like an asteroid whose trajectory you measure and you have certain instrument tolerances and me…

Sorry to break it to you, but all predictions are "made up". So much so you don't even know if you'll be dead or alive this time tomorrow.

The best we can do is have a Bayesian probability distribution about future outcomes.

I don't ever like holding a singular opinion on anything or converging on a single value for things like p(doom), but generally when I talk about things with people I try to represent my views as singular values or binary opinions. So I agree with you that p(doom) as a concept is silly in the same way as you saying you know you won't be dead in 24 hours is silly.

I can't speak for everyone, but I suspect most people who express a p(doom) haven't thought much about it so I agree with you there too. But I do personally put a decent amount of effort modelling out scenarios so I'll just speak to my own p(doom).

My p(doom) is high largely because I believe certain things about how AI capabilities will scale, how human incentives will shape certain outcomes, and our inability to slow progress in a field like AI. Basically:

- It's almost impossible to significantly slow AI progress without a doom-lite scenarios like nuclear war

- AI will scale to ASI in the coming decades

- Corruption is the norm, and democratic systems today are backed largely by the fact the public hold a monopoly over economic power - and all power derived from economic prosperity

- AI will replace 90% of productive human labour in the coming decades and erode the economic value of the humans, resulting over time in political corruption (resource curse)

- ASI unlocks all tech on the tech tree which primarily requires intelligence (which is likely to be the majority of tech)

- There is far more destructive/dangerous tech later on the tech tree than there is good. E.g., for every cancer cure there is 1,000s bioweapons and new torture techniques we can prompt Claude to create

- Even if humanity could agree on a workable definition of alignment it's a much harder problem than advancing capabilities, therefore capabilities will almost certainly progress faster than our ability to align advanced AI

- The field of alignment at it's limit resolves to basically just deciding between a god that follows the commands of a single or small group of humans (dangerous), or a god that has it's own moral considerations and does not listen to humans (dangerous).

So my p(doom) is largely a product of all of the above, and my general inability to find many neutral or good scenarios of ASI while landing on almost endless ways ASI is likely to go wrong regardless of whether you believe terminator-like scenarios. To be honest the only reason my personal p(doom) isn't higher than 99% is that I tend to weight the opinions of others whose opinions I value quite highly, and most people whose opinions I value have a slightly lower p(doom). I've rerated a lot recently because I was hoping some experts would turn out to be right about there being algorithmic limitations to ASI and things like "alignment by default". But I now think I've over weighted those opinions historically and a lot of the better arguments against a high p(doom) are either now dead in the water or have become increasingly hard to hold with any significant probability.

My best argument for against an AI doom scenario this century would be nuclear war, which I personally discard in my AI p(doom) probability. Allowing for other doom scenarios my AI p(doom) is probably closer to 90%.

Re: Who Aligns the Aligners?

#85
post #79
post #74

Earlier quoted context omitted.

From the point of view of the old people, exponentially large injections of improbability. This is also where it is important to understand the argument is parameterized. If your acceptance function is "a full human body living a full human life in the way I'm used to", then you're asking for a huge whackload of improbability to keep that body alive. If, on the other hand, you merely ask for "the minimum system that…

Quantum immortality is just an absurd quasi-religious pretense of physics that caters to those that fear death but feel embarrassed to participate in the traditional socially approved coping mechanisms.

I do not think death exists as a thing. We're shoving a lot of stuff into the concept. Pain, loss of information, of the one who dies. Fear of the unknown, what happens next. Death is a very volatile concept if you really try to pin it down. When you take all the layers off...you are left with nothing really.

But yes, the apparent loss of information seems to be true, for us at this technological level.

Re: Who Aligns the Aligners?

#86
post #68

Earlier quoted context omitted.

You're making a ridiculously improbable prediction that can never be falsified -- akin to beliefs in a vengeful God . Russell's Teapot is exactly the correct metaphor. The philosophical burden of proof lies on the person making empirically unfalsifiable claims. That is the point of Russell's Teapot.

> The philosophical burden of proof lies on the person making empirically unfalsifiable claims. You appear not to understand what a prediction is. "I believe X will likely happen because Y" is different from an assertion of fact that "X will happen". A prediction cannot even be empirically proved! You are simply suggesting that nobody is allowed to make a prediction, ever. Russell's Teapot, too, is not an argument th…

You can make whatever predictions you like. Don't expect other people to believe them and act on them unless you provide evidence.

Re: Who Aligns the Aligners?

#88
post #21

No one wants a slowdown. It's just a load of nonsense. It's something for the media to focus on, and it has no real impact. Everyone wants to rule the world and couldn't care less about anything else. People naively believe that this is even a serious issue.

>Everyone wants to rule the world and couldn't care less about anything else.

This might be scarily true. But we end up in the same problem, which is the fear of most people being wiped up.

Does it matter if the AI is doing it on its own or at the wishes of a small group of powerful people? For all intents and purposes they are both the same thing in the end, for most people anyway.

Re: Who Aligns the Aligners?

#89
post #5

> Nobody knows the answer and history is no guide, save that apocalyptic predictions about new technologies have, to date, all been wrong. If any past apocalyptic prediction had been true, you would not have been around to write this sentence and we would not have been around to read it. It’s not a very convincing argument when we cannot in principle observe the counterfactual.

But warnings (e.g. not apocalyptic but damaging or having any effects) have come true, by and large; people warned about the effects of pollution / co2 emissions / global warming over 150 years ago [0], to name a popular example.

I don't think there are many truly human created apocalyptic events out there besides global thermonuclear war, and I don't think the worst-case AI superintelligent killer robot scenario would be apocalyptic. Unless it somehow triggers global thermonuclear war, but then it becomes a chicken-or-egg question (and it'll probably be humans pushing the proverbial button, as nukes are '60s tech and (hopefully) airgapped etc).

[0] https://en.wikipedia.org/wiki/Svante_Arrhenius

Re: Who Aligns the Aligners?

#90

Earlier quoted context omitted.

We're in the middle of finding it out with climate change.

Middle? We're just starting . The air conditioning technology didn't hit its wall of fire yet. What we're going to do when that tech will cease to cool us? Every AC system (from refrigerators to air conditioners) has an upper limit, which is around 41 degrees Celsius for T1 class.

> What we're going to do when that tech will cease to cool us?

Move. This is already happening, with water shortages forcing people to move. Climate change will (likely) cause famines and many deaths, but the bigger effect will be mass migration.

Post reply on HN