Learning with Stochastic Guidance for Navigation

11/27/2018
by   Linhai Xie, et al.
0

Due to the sparse rewards and high degree of environment variation, reinforcement learning approaches such as Deep Deterministic Policy Gradient (DDPG) are plagued by issues of high variance when applied in complex real world environments. We present a new framework for overcoming these issues by incorporating a stochastic switch, allowing an agent to choose between high and low variance policies. The stochastic switch can be jointly trained with the original DDPG in the same framework. In this paper, we demonstrate the power of the framework in a navigation task, where the robot can dynamically choose to learn through exploration, or to use the output of a heuristic controller as guidance. Instead of starting from completely random moves, the navigation capability of a robot can be quickly bootstrapped by several simple independent controllers. The experimental results show that with the aid of stochastic guidance we are able to effectively and efficiently train DDPG navigation policies and achieve significantly better performance than state-of-the-art baselines models.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
08/01/2023

Target Search and Navigation in Heterogeneous Robot Systems with Deep Reinforcement Learning

Collaborative heterogeneous robot systems can greatly improve the effici...
research
06/17/2019

Robotic Navigation using Entropy-Based Exploration

Robotic navigation concerns the task in which a robot should be able to ...
research
10/26/2020

On Embodied Visual Navigation in Real Environments Through Habitat

Visual navigation models based on deep learning can learn effective poli...
research
02/13/2023

Improving robot navigation in crowded environments using intrinsic rewards

Autonomous navigation in crowded environments is an open problem with ma...
research
03/05/2023

Vision based Virtual Guidance for Navigation

This paper explores the impact of virtual guidance on mid-level represen...
research
07/11/2018

Learning Deployable Navigation Policies at Kilometer Scale from a Single Traversal

Model-free reinforcement learning has recently been shown to be effectiv...
research
11/21/2019

Accelerating Reinforcement Learning with Suboptimal Guidance

Reinforcement Learning in domains with sparse rewards is a difficult pro...

Please sign up or login with your details

Forgot password? Click here to reset