I don't think we can learn much from how an engine learns, but we certainly can learn from its results. For example, there's this interesting discussion: https://www.reddit.com/r/chess/comments/7ibzq4/stockfish_vs_... Because Alphazero did not learn from human games, it looks at the different pieces without attaching values like we do. It has no problems sacrificing a higher "valued" piece for the sake of its strateg…
Ask HN: Can we adapt AlphaZero's self-play technique for better human learning?
31–38 of 38 posts
Re: Ask HN: Can we adapt AlphaZero's self-play technique for better human learning?
#32Earlier quoted context omitted.
What are the winning conditions for math?
A smaller proof using fewer axioms or other proofs than the current state-of-the-art. Discovering new and "interesting" proofs. Don't ask me to define "interesting" in this context.
Re: Ask HN: Can we adapt AlphaZero's self-play technique for better human learning?
#33Earlier quoted context omitted.
What are the winning conditions for math?
A smaller proof using fewer axioms or other proofs than the current state-of-the-art. Discovering new and "interesting" proofs. Don't ask me to define "interesting" in this context.
Re: Ask HN: Can we adapt AlphaZero's self-play technique for better human learning?
#34Re: Ask HN: Can we adapt AlphaZero's self-play technique for better human learning?
#35Remember that AlphaZero played 44 million games of chess, whereas your average professional chess player has played somewhere on the order of 10,000-100,000. Self-play works, but rather slowly.
Re: Ask HN: Can we adapt AlphaZero's self-play technique for better human learning?
#36Earlier quoted context omitted.
>unfortunately I'm blanking on his name John Nash (supposedly) had this mindset? Is that who you're thinking about?
That wasn't who I had in mind, but thanks for sharing that example. I think the guy I'm thinking of is at Cornell and still alive. He also might actually be in CS instead of Math. I tried googling it but, unfortunately, "math professor who doesn't read papers" didn't come up with any results.
Re: Ask HN: Can we adapt AlphaZero's self-play technique for better human learning?
#37Re: Ask HN: Can we adapt AlphaZero's self-play technique for better human learning?
#38Earlier quoted context omitted.
>unfortunately I'm blanking on his name John Nash (supposedly) had this mindset? Is that who you're thinking about?
That wasn't who I had in mind, but thanks for sharing that example. I think the guy I'm thinking of is at Cornell and still alive. He also might actually be in CS instead of Math. I tried googling it but, unfortunately, "math professor who doesn't read papers" didn't come up with any results.