3 ms·
My mind went to Q learning.
by bitshiftfaced 3y ago
My mind went to Q learning.
- ilaksh 3y agoOk so maybe nothing to do with A*, but actually a way for GPT-powered models or agents to learn through automated reinforcement learning. Or something. I wonder if DeepMind is working on something similar also. If your hunch is right, this could lead to the type of self-improvement that scares people.
- Davidzheng 3y agoCould easily be both
- flibble 3y agoMine to Q of Star Trek.
- Izkata 3y agoBut were you thinking of Q, Q, or Q?
- deleted 3y ago[deleted]
- eli_gottlieb 3y agoMy mind went to some kind of Q-learning combined with something like a Monte Carlo Tree Search with some kind of A*-style heuristic to effectively combine Q-learning and with short-horizon planning.
- theGnuMe 3y agoThis was alpha-go and alpha-zero right?
- a-dub 3y agolikewise. i can already imagine a* being useful for efficiently solving basic algebra and proofs. it could form the basis of a generalized planning engine and that planning engine could potentially be dangerous given the inherent competitive reasoning behind any minmax style approach.