3 ms·
Reinforcement learning with unsupervised auxiliary tasks
- TrevorReznk 10y agoCan't wait to see how it will handle Starcraft 2.
- tener 10y ago> On Atari the agent now achieves on average 9x human performance Very impressive. I guess the human limit has to do with humans being limited about number of things to track at once? I wonder if they can apply this to optimizing the traffic lights in a big city.