3 ms·
I wonder if this is a tractable reinforcement learning problem. The objective is somewhat clearly defined-minimize spatial deviation over time from the input pa
by fromthestart 7y ago
I wonder if this is a tractable reinforcement learning problem. The objective is somewhat clearly defined-minimize spatial deviation over time from the input path.
- klodolph 7y agoIn order to train the model, wouldn’t you want a physics simulation, so you can train the model quickly? But if you had a physics simulation, aren’t there much easier methods to use than machine learning?