Mutual Information Maximization for Robust Plannable Representations

05/16/2020
by   Yiming Ding, et al.
34

Extending the capabilities of robotics to real-world complex, unstructured environments requires the need of developing better perception systems while maintaining low sample complexity. When dealing with high-dimensional state spaces, current methods are either model-free or model-based based on reconstruction objectives. The sample inefficiency of the former constitutes a major barrier for applying them to the real-world. The later, while they present low sample complexity, they learn latent spaces that need to reconstruct every single detail of the scene. In real environments, the task typically just represents a small fraction of the scene. Reconstruction objectives suffer in such scenarios as they capture all the unnecessary components. In this work, we present MIRO, an information theoretic representational learning algorithm for model-based reinforcement learning. We design a latent space that maximizes the mutual information with the future information while being able to capture all the information needed for planning. We show that our approach is more robust than reconstruction objectives in the presence of distractors and cluttered scenes

READ FULL TEXT
research
12/09/2019

Learning Latent State Spaces for Planning through Reward Prediction

Model-based reinforcement learning methods typically learn models for hi...
research
07/17/2021

High-Accuracy Model-Based Reinforcement Learning, a Survey

Deep reinforcement learning has shown remarkable success in the past few...
research
06/19/2023

Learning Models of Adversarial Agent Behavior under Partial Observability

The need for opponent modeling and tracking arises in several real-world...
research
06/14/2021

Temporal Predictive Coding For Model-Based Planning In Latent Space

High-dimensional observations are a major challenge in the application o...
research
05/08/2020

Efficient Reconstruction of Stochastic Pedigrees

We introduce a new algorithm called Rec-Gen for reconstructing the gene...
research
03/22/2021

Volumetric Objectives for Multi-Robot Exploration of Three-Dimensional Environments

Volumetric objectives for exploration and perception tasks seek to captu...
research
03/03/2022

Compressed Predictive Information Coding

Unsupervised learning plays an important role in many fields, such as ar...

Please sign up or login with your details

Forgot password? Click here to reset