3 ms·
Sounds like what you're talking about is decision theory. Or, when mixed with learning: reinforcement learning. The monte carlo AIXI approximation would be an e
by badmofo666 13y ago
Sounds like what you're talking about is decision theory. Or, when mixed with learning: reinforcement learning. The monte carlo AIXI approximation would be an example. But, it really only works well on small toy problems (I.e. problems where the agent has only a small number of available actions it can perform).
- wlievens 13y agoThing is, humans come up with new "actions" all the time. If you drop that abstraction, and reduce our actions to "controlling information flow in our bodies" then the action space becomes unfathomably huge.