Document results for

Cite This For Me: Harvard, APA, MLA Reference Generato

7 results · Page 1 of 1

PDF

[2310.16029] Finetuning Offline World Models in the Real World

In this work, we seek to get the best of both worlds: we consider the problem of pretraining a world model with offline data collected on a real robot, and then finetuning the model on online data collected by planning with the learned model.

arxiv.org
PDF

Efficient Reinforcement Learning by Guiding World Models with ...

The work addresses the challenge of improving sample efficiency in reinforcement learning (RL) by leveraging non-curated offline datasets which are reward-free, of mixed quality, and collected from multiple embodiments during online reinforcement learning.

openreview.net
PDF

Investigating Online RL in World Models - OpenReview

We propose an algorithm and a data curation method that addresses both of these concerns by demonstrating that effective full-length rollout training is possible without hand-crafted penalties by treating each member of the world model ensemble as a level in the Unsupervised Environment Design (UED) framework.

openreview.net

Find related books on Amazon

As an Amazon Associate, we earn from qualifying purchases.

Cite This For Me: Harvard, APA, MLA Reference Generato

Cite This For Me: Harvard, APA, MLA Reference Generato book

Cite This For Me: Harvard, APA, MLA Reference Generato handbook