Define canonical, but the encyclopedia is a great start: https://en.wikipedia.org/wiki/Existential_risk_from_artifici... Fascinating that even Turing had considered the possibility.
Is there a canonical source for “the argument for AGI ruin” somewhere
21–30 of 81 posts
Re: Is there a canonical source for “the argument for AGI ruin” somewhere
#22Earlier quoted context omitted.
Of course, though I surmise that he’s using “canonical” as a metaphor for “best argued” or “most robust proof.” For example though The 1963 paper titled "Is Justified True Belief Knowledge?", is the canonical source of the Gettier Problem That was less than 100 years ago, so definitely not classical. I doubt the time period is relevant to whether you’ve made the best/canonical argument on a topic.
For example, if I asked for the canonical argument for the existence of god, I'd probably get a wide variety of answers proposing essays from various theologicians throughout history, like Thomas Aquinas or René Descartes. Since there are a large number of them, no single one of them can be the canonical source. Perhaps one could link the Wikipedia page ( https://en.wikipedia.org/wiki/Existence_of_God ) which summari…
Re: Is there a canonical source for “the argument for AGI ruin” somewhere
#23I admit I’m completely out of my depth when it comes to this field — I don’t even typically care about AI — but Eliezer’s response looks really bad to anyone in research. Asking for a citable and thorough written argument is as basic a requirement as it gets. To repudiate that request with “everyone has a different objection” is nearly unthinkable. And to David Chalmers no less! I think podcast hosts, tweet authors,…
Re: Is there a canonical source for “the argument for AGI ruin” somewhere
#24Re: Is there a canonical source for “the argument for AGI ruin” somewhere
#25https://www.alignmentforum.org/posts/pRkFkzwKZ2zfa3R6H/witho...
Personally, I'm encouraged by the emergence of chain-of-thought prompting for LLMs. Machine learning models have a reputation for being opaque and impossible to interpret. But right now, the best way to get LLMs to perform more complex logical reasoning is to make them write out that reasoning, a mechanism which happens to have built-in interpretability. Perhaps future advances in reasoning will involve more opaque internal states, but it seems plausible to me that the goals of 'be good at human-like reasoning' and 'be able to explain that reasoning (in the way humans do)' will continue to be well-aligned in the future. There would still be the possibility of the AI learning to be deceptive when explaining itself, but it would be much more difficult.
Re: Is there a canonical source for “the argument for AGI ruin” somewhere
#26Is there a universally accepted definition of AGI?
For example self-awareness might be part of a definition of intelligence. But an AI can lie about it’s self awareness.
Re: Is there a canonical source for “the argument for AGI ruin” somewhere
#27Re: Is there a canonical source for “the argument for AGI ruin” somewhere
#28Which is not to downplay AGI risk per se- it's a powerful tool, and powerful anything can be dangerous. But the uniquely foomy paperclip maximizer fear? That's a cultural attractor from the bay area through and through.
Re: Is there a canonical source for “the argument for AGI ruin” somewhere
#29I admit I’m completely out of my depth when it comes to this field — I don’t even typically care about AI — but Eliezer’s response looks really bad to anyone in research. Asking for a citable and thorough written argument is as basic a requirement as it gets. To repudiate that request with “everyone has a different objection” is nearly unthinkable. And to David Chalmers no less! I think podcast hosts, tweet authors,…
What's the academically supported description/ argument or whatever for why nuclear weapons are dangerous for the world if they exist? Obviously they are dangerous for the world because we might kill everyone if someone makes a mistake or gets mad and shoots one first. The danger of nukes is more obvious than why AI is dangerous.
The Wikipedia article on mutual assured destruction (https://en.wikipedia.org/wiki/Mutual_assured_destruction) traces the origin and evolution of exactly that analysis, including even pre-atom bomb precursors based on earlier weapons that would hypothetically make war "too terrible to ever happen again".
Re: Is there a canonical source for “the argument for AGI ruin” somewhere
#30I admit I’m completely out of my depth when it comes to this field — I don’t even typically care about AI — but Eliezer’s response looks really bad to anyone in research. Asking for a citable and thorough written argument is as basic a requirement as it gets. To repudiate that request with “everyone has a different objection” is nearly unthinkable. And to David Chalmers no less! I think podcast hosts, tweet authors,…
Presumably, Yudkowksy could point to any of several of his own publications, like this one: https://intelligence.org/files/AIPosNegFactor.pdf