3 ms·
True, but the next logical step is to put these models in a standard reinforcement learning environment. See: https://sites.research.google/palm-saycan https://
by casebash 4y ago
True, but the next logical step is to put these models in a standard reinforcement learning environment. See: https://sites.research.google/palm-saycan https://sites.research.google/palm-saycan
(I mean like with a proper world model and not just RLHF which they are already doing).