Alert button

Planning from Pixels using Inverse Dynamics Models

Dec 04, 2020
Keiran Paster, Sheila A. McIlraith, Jimmy Ba

Figure 1 for Planning from Pixels using Inverse Dynamics Models
Figure 2 for Planning from Pixels using Inverse Dynamics Models
Figure 3 for Planning from Pixels using Inverse Dynamics Models
Figure 4 for Planning from Pixels using Inverse Dynamics Models

Share this with someone who'll enjoy it:

Learning task-agnostic dynamics models in high-dimensional observation spaces can be challenging for model-based RL agents. We propose a novel way to learn latent world models by learning to predict sequences of future actions conditioned on task completion. These task-conditioned models adaptively focus modeling capacity on task-relevant dynamics, while simultaneously serving as an effective heuristic for planning with sparse rewards. We evaluate our method on challenging visual goal completion tasks and show a substantial increase in performance compared to prior model-free approaches.

* 9 pages, 4 figures  
View paper onarxiv iconopen_review iconOpenReview

Share this with someone who'll enjoy it: