Planning from Pixels using Inverse Dynamics Models (2012.02419v1)

Published 4 Dec 2020 in cs.LG and cs.AI

Abstract: Learning task-agnostic dynamics models in high-dimensional observation spaces can be challenging for model-based RL agents. We propose a novel way to learn latent world models by learning to predict sequences of future actions conditioned on task completion. These task-conditioned models adaptively focus modeling capacity on task-relevant dynamics, while simultaneously serving as an effective heuristic for planning with sparse rewards. We evaluate our method on challenging visual goal completion tasks and show a substantial increase in performance compared to prior model-free approaches.

Citations (39)

View on Semantic Scholar

Summary

We haven't generated a summary for this paper yet.

Summarize Now

Planning from Pixels using Inverse Dynamics Models (2012.02419v1)

Summary

Related Papers