Temporal-adaptive Hierarchical Reinforcement Learning

02/06/2020
by   Wen-Ji Zhou, et al.
0

Hierarchical reinforcement learning (HRL) helps address large-scale and sparse reward issues in reinforcement learning. In HRL, the policy model has an inner representation structured in levels. With this structure, the reinforcement learning task is expected to be decomposed into corresponding levels with sub-tasks, and thus the learning can be more efficient. In HRL, although it is intuitive that a high-level policy only needs to make macro decisions in a low frequency, the exact frequency is hard to be simply determined. Previous HRL approaches often employed a fixed-time skip strategy or learn a terminal condition without taking account of the context, which, however, not only requires manual adjustments but also sacrifices some decision granularity. In this paper, we propose the temporal-adaptive hierarchical policy learning (TEMPLE) structure, which uses a temporal gate to adaptively control the high-level policy decision frequency. We train the TEMPLE structure with PPO and test its performance in a range of environments including 2-D rooms, Mujoco tasks, and Atari games. The results show that the TEMPLE structure can lead to improved performance in these environments with a sequential adaptive high-level control.

READ FULL TEXT
research
01/24/2022

Adversarially Guided Subgoal Generation for Hierarchical Reinforcement Learning

Hierarchical reinforcement learning (HRL) proposes to solve difficult ta...
research
10/11/2022

DHRL: A Graph-Based Approach for Long-Horizon and Sparse Hierarchical Reinforcement Learning

Hierarchical Reinforcement Learning (HRL) has made notable progress in c...
research
05/18/2021

Adaptive ABAC Policy Learning: A Reinforcement Learning Approach

With rapid advances in computing systems, there is an increasing demand ...
research
12/26/2022

Learning Generalizable Representations for Reinforcement Learning via Adaptive Meta-learner of Behavioral Similarities

How to learn an effective reinforcement learning-based model for control...
research
06/25/2019

Reinforcement Learning with Competitive Ensembles of Information-Constrained Primitives

Reinforcement learning agents that operate in diverse and complex enviro...
research
12/08/2015

Reinforcement Control with Hierarchical Backpropagated Adaptive Critics

Present incremental learning methods are limited in the ability to achie...
research
03/11/2021

Adapting User Interfaces with Model-based Reinforcement Learning

Adapting an interface requires taking into account both the positive and...

Please sign up or login with your details

Forgot password? Click here to reset