4 ms·Reinforcement Learning As One Big Sequence Modeling Problem14 points by optimalsolver 5y agoT-A 5y agohttps://arxiv.org/abs/2106.01345 https://arxiv.org/abs/2106.01345