Reinforcement co-Learning of Deep and Spiking Neural Networks for Energy-Efficient Mapless Navigation with Neuromorphic Hardware

03/02/2020
by   Guangzhi Tang, et al.
0

Energy-efficient mapless navigation is crucial for mobile robots as they explore unknown environments with limited on-board resources. Although the recent deep reinforcement learning (DRL) approaches have been successfully applied to navigation, their high energy consumption limits their use in many robotic applications. Here, we propose a neuromorphic approach that combines the energy-efficiency of spiking neural networks with the optimality of DRL to learn control policies for mapless navigation. Our hybrid framework, Spiking deep deterministic policy gradient (SDDPG), consists of a spiking actor network (SAN) and a deep critic network, where the two networks were trained jointly using gradient descent. The trained SAN was deployed on Intel's Loihi neuromorphic processor. The co-learning enabled synergistic information exchange between the two networks, allowing them to overcome each other's limitations through a shared representation learning. When validated on both simulated and real-world complex environments, our method on Loihi not only consumed 75 times less energy per inference as compared to DDPG on Jetson TX2, but also had a higher rate of successfully navigating to the goal which ranged by 1% to 4.2%, depending on the forward-propagation timestep size. These results reinforce our ongoing effort to design brain-inspired algorithms for controlling autonomous robots with neuromorphic hardware.

READ FULL TEXT

page 1

page 4

page 5

page 6

page 7

research
10/19/2020

Deep Reinforcement Learning with Population-Coded Spiking Neural Network for Continuous Control

The energy-efficient control of mobile robots is crucial as the complexi...
research
10/05/2022

Neuro-Planner: A 3D Visual Navigation Method for MAV with Depth Camera based on Neuromorphic Reinforcement Learning

Traditional visual navigation methods of micro aerial vehicle (MAV) usua...
research
03/06/2019

Spiking Neural Network on Neuromorphic Hardware for Energy-Efficient Unidimensional SLAM

Energy-efficient simultaneous localization and mapping (SLAM) is crucial...
research
03/26/2022

A Novel Neuromorphic Processors Realization of Spiking Deep Reinforcement Learning for Portfolio Management

The process of continuously reallocating funds into financial assets, ai...
research
03/06/2020

Evolved Neuromorphic Control for High Speed Divergence-based Landings of MAVs

Flying insects are capable of vision-based navigation in cluttered envir...
research
05/04/2021

Simplified Klinokinesis using Spiking Neural Networks for Resource-Constrained Navigation on the Neuromorphic Processor Loihi

C. elegans shows chemotaxis using klinokinesis where the worm senses the...
research
03/05/2021

A Dual-Memory Architecture for Reinforcement Learning on Neuromorphic Platforms

Reinforcement learning (RL) is a foundation of learning in biological sy...

Please sign up or login with your details

Forgot password? Click here to reset