This actually brings to mind a conversation I've been having with my buddies recently. Has there ever been a case of harm from AI misalignment? Not some kind of speculative art exhibit, but real harm?
Misalignment Museum
21–30 of 64 posts
Re: Misalignment Museum
#22Re: Misalignment Museum
#23Wow, didn't expect my https://www.niche-museums.com/ website to show up on Hacker News! It's actually running on a templated instance of Datasette - try the "Use my location" button on the homepage to see tiny museums near you that I've been to.
May I suggest two additions? Kassel, Germany has two weird ones: A museum for death culture (Sepulkralkulturmuseum, https://www.sepulkralmuseum.de/), and a museum of wallpaper (https://www.tapeten.museum-kassel.de/) :)
Re: Misalignment Museum
#24[flagged]
Re: Misalignment Museum
#25Wow, didn't expect my https://www.niche-museums.com/ website to show up on Hacker News! It's actually running on a templated instance of Datasette - try the "Use my location" button on the homepage to see tiny museums near you that I've been to.
Is there a way to add entries to this collection, or do you only list museums you have been personally been at? I ask because I tried look for listings of places I know well (e.g. Turin, Italy or Budapest, Hungary) and found that the closest entry was hundreds of miles away. Turin has, for example, this: http://www.museodellafrutta.it/en/ which is pretty niche, imho. Budapest has a chocolate museum: https://www.csoko…
The one I would be recommending (to the niche museums site, and also to everyone who is around) is the Yokohama Coast Guard museum.
It has a really unassuming name, I only went in because I got caught out in the rain without an umbrella in Minatomirai. Dare I say it is the best museum I have ever visited. It is organised around a single curious event: In 2001 a Japanese Coast Guard vessel encountered a suspiciously behaving fishing travel. They wanted to board them when the fishing vessel took off at high speed and started shooting at them. During the pursuit they even seen the crew wield shoulder mounter missile launchers.
Turns out it was a North Korean spy ship on a mission to raise funds by smuggling drugs to Japan. The incident ended by the Korean crew scuttling their vessel. The coast guard has raised the sunk ship and built this museum around it.
Website of the museum: https://jcgmuseum.jp/en/
More info on the event: https://en.wikipedia.org/wiki/Battle_of_Amami-%C5%8Cshima
Re: Misalignment Museum
#26Earlier quoted context omitted.
The problem with that is - until we have "true AGI" about which everybody agree it's "true AGI" - you can always dismiss deaths caused by software as "just a bug, not an AI safety problem". For example when Tesla autopilot kills someone (which has happened). https://impakter.com/tesla-autopilot-crashes-with-at-least-a...
I guess maybe one could make that argument, but the majority of AI alignment concerns seem more along the lines of skynet level scenarios. The paperclip optimizer could be called a "bug" I guess. But the comparison to Tesla autopilot is interesting. First, is anyone calling that an AI? Second, when an objective baseline comparison to human drivers is possible, shouldn't that determine whether the AI is a net benefit…
Humanity doesn't hate polar bears. We just love burning oil more than we love them. Polar bears die as an unintended side effect. We might even be sad about that. We might even keep a few alive in ZOOs. But we won't change our whole economy to save a few cute bears.
Of course these kinds of threats will start to appear only when it's smart enough. The main problem with AGI is that we probably won't know until it's too late, because of how fast this technology develops once it can improve itself.
Can I ask you another question - what do you think will happen?
1. it won't get smarter than us 2. it will care about us 3. we will somehow keep it in check despite the fact it's smarter than us
Cause it seems to me that most people still intuitively think (1) because it's "too sci-fi". And even if they became persuaded that (1) is no longer certain - they didn't updated the rest of their beliefs with that new information, they are still believing AI is safe like they did when (1) was assumed true because they haven't updated the cache, or maybe they don't even realize there's dependency somewhere in their train of thought that needs to be updated.
This is how living in a world undergoing singularity will be like, BTW - you can't think one thought to the end without realizing the assumptions might have changed since the last time you thought about it. So you go down and realize the assumptions there are also changing. And so on.
Re: Misalignment Museum
#27This actually brings to mind a conversation I've been having with my buddies recently. Has there ever been a case of harm from AI misalignment? Not some kind of speculative art exhibit, but real harm?
Re: Misalignment Museum
#28Wow, didn't expect my https://www.niche-museums.com/ website to show up on Hacker News! It's actually running on a templated instance of Datasette - try the "Use my location" button on the homepage to see tiny museums near you that I've been to.
Is there a way to add entries to this collection, or do you only list museums you have been personally been at? I ask because I tried look for listings of places I know well (e.g. Turin, Italy or Budapest, Hungary) and found that the closest entry was hundreds of miles away. Turin has, for example, this: http://www.museodellafrutta.it/en/ which is pretty niche, imho. Budapest has a chocolate museum: https://www.csoko…
My favorites are the Clockarium, Museum of the Art Deco Ceramic Clock. (https://www.clockarium.org/ Check out the video, its is magnificent.)
And The Sewer Museum: Experience an authentic sewer, stroll along the Senne and discover the little-known but ever so important profession of a sewage worker. Descend deep into the bowels of the city for this unique experience! https://sewermuseum.brussels/
Re: Misalignment Museum
#29Wow, didn't expect my https://www.niche-museums.com/ website to show up on Hacker News! It's actually running on a templated instance of Datasette - try the "Use my location" button on the homepage to see tiny museums near you that I've been to.
Re: Misalignment Museum
#30Earlier quoted context omitted.
The problem with that is - until we have "true AGI" about which everybody agree it's "true AGI" - you can always dismiss deaths caused by software as "just a bug, not an AI safety problem". For example when Tesla autopilot kills someone (which has happened). https://impakter.com/tesla-autopilot-crashes-with-at-least-a...
I guess maybe one could make that argument, but the majority of AI alignment concerns seem more along the lines of skynet level scenarios. The paperclip optimizer could be called a "bug" I guess. But the comparison to Tesla autopilot is interesting. First, is anyone calling that an AI? Second, when an objective baseline comparison to human drivers is possible, shouldn't that determine whether the AI is a net benefit…
A software bug is generally what we call the situation where software does exactly what its creator asked it to do, but that thing is not a thing its creator actually intended or wanted. The creator did not correctly express their request, and/or did not properly think through all the effects of carrying it out.
Think about people giving each other instructions and making rules for each other as we normally do day to day in English. English is not a very precise language for expressing what we actually want to happen, and also humans are not very good at rigorously specifying what they want to happen, relying instead on assumed implicit shared understanding; these assumptions lead to much misery in human/human interactions, never mind human/computer. Worse, humans are not very good at actually knowing either what they want the world to be like or how to make that happen. With the best of intentions, we make rules and set policies, intending that the world become better for it, and for every instruction we give, rule we set, policy we make, we invariably end up with some unintended consequences. This is the human condition: every day, we try to make the world a little better, but in the end things turn out like they always do. General human communication relies on shared values, but our values are not actually universally shared, we don't know how to even begin rigorously expressing our values, and much of the time we can't agree on what they are or even properly explain our own values to ourselves - all those fuzzy open questions in philosophy arise from this.
When we interact with a general AI, we are programming a computer system, in English. On top of all the usual problems, the computer system lacks our shared understanding, because not only do we not know how to impart it, we don't even really know what to impart. It is not aligned to human values. It can't be: we can't even achieve alignment with each other, never mind a lump of silicon. So miscommunication is inevitable. Worse, the recent direction in AI has been to throw away any attempt to actually express what behaviours we want explicitly, and instead just throw the entire contents of the internet at a giant statistical model and hope the correlations it makes are somehow useful to us. The honestly surprising thing is that this even works to any extent at all. But a few minutes' interaction quickly assures us that the resulting systems react quite unpredictably to our input.
We will ask for things, the AI will do exactly what we asked for, and we will find that what we literally asked for is not actually what we want: the system that is the combination of our request with the AI will contain bugs.
The amount of resulting harm is determined by what it is the software is controlling and how much time we have to react to the unintended consequences. The AI doomer claim is that, unless we do better at the alignment problem, as our software gets faster and we give it control over more stuff, the inevitable bugs will cause bad things to happen faster than we can react to prevent harm, and the consequences will be worse than we can tolerate; worse, as we link everything together, we might not even realise we are working indirectly with safety critical systems until they do whatever it is we told them to but did not mean.
The solution should be obvious: don't put software in control of devices that interact with the real world in ways that could cause serious harm without first rigorously proving that no combination of inputs can result in behaviours that make the situation worse instead of better. Traditional engineering sounds expensive, hard and tedious, and it is, and it is not very shiny or sexy, but we can do it, and do do it in situations where serious harm will otherwise result, like aviation or (most of) the automotive industry. Include the fuzzy ill-conditioned statistics software, by all means, but don't wire it directly to the controls - make it an input to a traditionally engineered and well understood system, treated like any other noisy and potentially broken input, with the system as a whole rigorously designed to produce safe outputs when it can and to safely shut down when it cannot.
Surely the AI doomers are overstating the risk of doom - surely no-one working in safety critical systems would do things any other way? "Has there ever been a case of harm form AI misalignment?" - this is the real thrust of that question. What sort of idiot would wire an unintended consequence generator directly to anything that might harm or hurt?
Tesla autopilot is interesting precisely as a current ongoing real-world example of the new-style fuzzy black box tech being wired directly to several tons of trundling metal, harm to property and life resulting from unintended behaviours, and instead of putting things on hold pending fault analysis and a more rigorous approach, we just double down on throwing more data at the fuzzy black box and hoping the fact that we can't reproduce the last bug with a few quick tests means it's gone away.