Efficient Reinforcement Learning by Guiding World Models with ...
The work addresses the challenge of improving sample efficiency in reinforcement learning (RL) by leveraging non-curated offline datasets which are reward-free, of mixed quality, and collected from multiple embodiments during online reinforcement learning.
Source: openreview.net