Live data from Hacker News

Deep Mind Playing Montezuma's Revenge with Intrinsic Motivation [video]

youtube.com

1–10 of 46 posts

Re: Deep Mind Playing Montezuma's Revenge with Intrinsic Motivation [video]

#2
If you end up blocked by a popup blocker, or just don't feel like reading fluff, or like me you've got the browser locked down too tightly for their integrated video player to work, the paper is at https://arxiv.org/pdf/1606.01868v1.pdf and the video mentioned is (probably, since I didn't see the original) at https://www.youtube.com/watch?v=0yI2wJ6F8r0 .

Re: Deep Mind Playing Montezuma's Revenge with Intrinsic Motivation [video]

#3
It might be a interesting future if computers become powerful enough to be able to run this kind of IA. Then games companies might be able to use directly a generic IA like this instead of developing specific ones. But I suppose it will be hard to do and won't happen (if it does) before a long time.

Re: Deep Mind Playing Montezuma's Revenge with Intrinsic Motivation [video]

#6
post #5

Is it really true that Montezuma's Revenge is more challenging for DeepMind than Go, as they mention in the article?

Isn't DeepMind just an umbrella term like Watson? I would guess that AlphaGo and the software that learns games like this are distinct.

Re: Deep Mind Playing Montezuma's Revenge with Intrinsic Motivation [video]

#7
post #6
post #5

Is it really true that Montezuma's Revenge is more challenging for DeepMind than Go, as they mention in the article?

Isn't DeepMind just an umbrella term like Watson? I would guess that AlphaGo and the software that learns games like this are distinct.

DeepMind is the name of an AI company that was acquired by Google.

https://deepmind.com/

AlphaGo is the program that beat Go. These game playing algorithms are variations of Deep Q Learning Reinforcement Learning algos.

Re: Deep Mind Playing Montezuma's Revenge with Intrinsic Motivation [video]

#9
post #6

Earlier quoted context omitted.

Isn't DeepMind just an umbrella term like Watson? I would guess that AlphaGo and the software that learns games like this are distinct.

DeepMind is the name of an AI company that was acquired by Google. https://deepmind.com/ AlphaGo is the program that beat Go. These game playing algorithms are variations of Deep Q Learning Reinforcement Learning algos.

And here's the related paper, "Unifying Count-Based Exploration and Intrinsic Motivation": https://arxiv.org/abs/1606.01868

Tech journalism isn't at its best when it can't distinguish between WOPR and a team of researchers.

Post reply on HN