3 ms·Decaying Evidence and Contextual Bandits – Bayesian Reinforcement Learning2 points by _eigenfoo 7y ago