4 ms·
Super cool. How long did this take to make? I kind of wonder if there is some nice analogy to be made here wrt. Kelly Betting vs Bayesian RL. As in, some versi
by algon33 6y ago
Super cool. How long did this take to make?
I kind of wonder if there is some nice analogy to be made here wrt. Kelly Betting vs Bayesian RL. As in, some version of maximising log reward will have higher median performance than Bayesian RL even though on average Bayesian RL is better. By analogy, the discprepancy should come from Bayesian RL doing vastly better in some unlikely string of world trajectories.
- gwern 6y agohttps://www.reddit.com/r/MachineLearning/comments/jg475u/r_a_bayesian_perspective_on_qlearning/g9o8xpy/ https://www.reddit.com/r/MachineLearning/comments/jg475u/r_a...