Emergent Tool Use from Multi-Agent Interaction
61–64 of 64 posts
Re: Emergent Tool Use from Multi-Agent Interaction
#62Earlier quoted context omitted.
Good catch! Will update the post to be explicit that there are many pre-existing awesome results in this vein.
Any possibility of releasing the simulation environment? Looks quite cool!
Environment Generation: https://github.com/openai/multi-agent-emergence-environments
Worldgen: https://github.com/openai/mujoco-worldgen
Is this what you were looking for?
Re: Emergent Tool Use from Multi-Agent Interaction
#63Re: Emergent Tool Use from Multi-Agent Interaction
#64Earlier quoted context omitted.
By the same token, I’m extremely suspicious of the idea that such a sufficiently complex AGI could also be dumb enough to optimize for paper clip production at the expense of all life on earth (or w/e example).
...and many would say that’s because us humans are bad at imagining optimizing agents without anthropomorphizing them. This is a reasonable, even typical suspicion that many people share! The best explanation I know of why it’s unfortunately wrong is by Robert Miles in a video, but if you prefer a more thorough treatment, you could also read about “instrumental convergence” directly. If you find a flaw in this idea,…
If you assume an AGI is incapable of asking “why” about its terminal goal, you have to assume it’s incapable of asking “why” in any context. Miles’ AGI has no power of metacognition, but is still somehow able to reprogram itself. This really isn’t compatible with “general intelligence” or the powers that get ascribed to imaginary AGIs.
I’m certainly no expert, but I expect there will turn out to be something like the idea of Turing-completeness for AI. Just like any general computing machine is a computer, any true AGI will be sapient. You can’t just arbitrarily pluck a part out, like “it can’t reason about its objective”, and expect it to still function as an AGI, just like you can’t say “it’s Turing complete, except it can’t do any kind of conditional branching.” EDIT better example: “it’s Turing complete, but it can’t do bubble sort.”
This intuition may be wrong, but it’s just as much as assumption as Miles’ argument.
I’m also not ascribing morality to it: we have our share of psychopaths, and intelligence doesn’t imply empathy. AGI may very well be dangerous, just probably not the “mindlessly make paperclips” kind.