3 ms·
Tree search is also used during play. In the paper, they pit the pure neural net against other versions of the algorithm -- it ends up slightly worse than the
by panic 9y ago
Tree search is also used during play. In the paper, they pit the pure neural net against other versions of the algorithm -- it ends up slightly worse than the version that played Fan Hui, at about 3000 ELO.
- AlexCoventry 9y agoOh, so it's just not using rollouts to estimate the board position? Thanks for the clarification.
- mrec 9y agoIt doesn't use rollouts at all: > AlphaGo Zero does not use “rollouts” - fast, random games used by other Go programs to predict which player will win from the current board position. Instead, it relies on its high quality neural networks to evaluate positions.