3 ms·
That might work; the concern is that this would take it too far off of the task it was trained on. That is, if it doesn't have a lot of experience being signifi
by clickok 11y ago
That might work; the concern is that this would take it too far off of the task it was trained on.
That is, if it doesn't have a lot of experience being significantly down, then it won't play nearly as well when trying to catch up-- but that doesn't matter in even games because it never gets that far behind.
You're right that it would be interesting to see, though-- we need to get better at understanding these sorts of systems, at least until they can start optimizing themselves.
Alternatively, we might train a different agent (OmegaGo?) to try to win by the largest margin possible-- if it works as well as AlphaGo, then that might give us some more insight into how strong both programs are.