Earlier quoted context omitted.
You're thinking too much like a human. No, you're thinking too much like someone who understands chess but not Go. Suppose that we built an opening book with a trillion reasonable courses of action on it. Each one analyzed well. As you have discovered, you will only go a few ply into the game. And all of the positions that you will be directed towards will have only a small edge. Instead put a tiny fraction of the co…
> Instead put a tiny fraction of the computing power necessary to build this book into self-training. Do you believe that AlphaZero could continue to improve dramatically with another 6-months of training? Or if it can improve at another 6-months after that? At some point, the network will reach a local maximum, and it will be unable to improve beyond that. Characterizing AlphaZero's moves through big-data analysis i…
This only works if AlphaZero never retrains on previous games.