4 ms·End-to-End Robotic Reinforcement Learning Without Reward Engineering1 points by ChankeyPathak 7y agoChankeyPathak 7y agoLink to paper: https://arxiv.org/abs/1904.07854 https://arxiv.org/abs/1904.07854