True, but the next logical step is to put these models in a standard reinforcement learning environment. See: https://sites.research.google/palm-saycan
(I mean like with a proper world model and not just RLHF which they are already doing).
(I mean like with a proper world model and not just RLHF which they are already doing).
No comments yet.