2 ms·They can always finetune using RL later. Superversied training was the first step at making AlphaGo work.by dontreact 9y agoThey can always finetune using RL later. Superversied training was the first step at making AlphaGo work.