Toward Robust Long Range Policy Transfer

03/04/2021
by   Wei-Cheng Tseng, et al.
4

Humans can master a new task within a few trials by drawing upon skills acquired through prior experience. To mimic this capability, hierarchical models combining primitive policies learned from prior tasks have been proposed. However, these methods fall short comparing to the human's range of transferability. We propose a method, which leverages the hierarchical structure to train the combination function and adapt the set of diverse primitive polices alternatively, to efficiently produce a range of complex behaviors on challenging new tasks. We also design two regularization terms to improve the diversity and utilization rate of the primitives in the pre-training phase. We demonstrate that our method outperforms other recent policy transfer methods by combining and adapting these reusable primitives in tasks with continuous action space. The experiment results further show that our approach provides a broader transferring range. The ablation study also shows the regularization terms are critical for long range policy transfer. Finally, we show that our method consistently outperforms other methods when the quality of the primitives varies.

READ FULL TEXT

page 3

page 4

page 6

research
05/23/2019

MCP: Learning Composable Hierarchical Control with Multiplicative Compositional Policies

Humans are able to perform a myriad of sophisticated tasks by drawing up...
research
10/07/2021

Augmenting Reinforcement Learning with Behavior Primitives for Diverse Manipulation Tasks

Realistic manipulation tasks require a robot to interact with an environ...
research
10/25/2021

Learning Insertion Primitives with Discrete-Continuous Hybrid Action Space for Robotic Assembly Tasks

This paper introduces a discrete-continuous action space to learn insert...
research
03/03/2020

Hierarchically Decoupled Imitation for Morphological Transfer

Learning long-range behaviors on complex high-dimensional agents is a fu...
research
06/25/2019

Reinforcement Learning with Competitive Ensembles of Information-Constrained Primitives

Reinforcement learning agents that operate in diverse and complex enviro...
research
11/28/2018

Neural probabilistic motor primitives for humanoid control

We focus on the problem of learning a single motor module that can flexi...
research
06/24/2021

Towards Exploiting Geometry and Time for Fast Off-Distribution Adaptation in Multi-Task Robot Learning

We explore possible methods for multi-task transfer learning which seek ...

Please sign up or login with your details

Forgot password? Click here to reset