4 ms·Those were trained on human play. This had to figure it out from scratch.by breakyerself 2y agoThose were trained on human play. This had to figure it out from scratch.camel-cdr 2y agoAh, is this full RL? I was reading something about LLMs earlier and was thinking that LLMs could probably write a simple case based script for controlling a player, that could accive a decent success rate.danijar 2y agoYes, it's RL from scratch and sparse rewards