Reinforcement learning reward function in unmanned aerial vehicle control tasks

03/20/2022
by   Mikhail S. Tovarnov, et al.
0

This paper presents a new reward function that can be used for deep reinforcement learning in unmanned aerial vehicle (UAV) control and navigation problems. The reward function is based on the construction and estimation of the time of simplified trajectories to the target, which are third-order Bezier curves. This reward function can be applied unchanged to solve problems in both two-dimensional and three-dimensional virtual environments. The effectiveness of the reward function was tested in a newly developed virtual environment, namely, a simplified two-dimensional environment describing the dynamics of UAV control and flight, taking into account the forces of thrust, inertia, gravity, and aerodynamic drag. In this formulation, three tasks of UAV control and navigation were successfully solved: UAV flight to a given point in space, avoidance of interception by another UAV, and organization of interception of one UAV by another. The three most relevant modern deep reinforcement learning algorithms, Soft actor-critic, Deep Deterministic Policy Gradient, and Twin Delayed Deep Deterministic Policy Gradient were used. All three algorithms performed well, indicating the effectiveness of the selected reward function.

READ FULL TEXT
research
09/24/2020

Motion Planning by Reinforcement Learning for an Unmanned Aerial Vehicle in Virtual Open Space with Static Obstacles

In this study, we applied reinforcement learning based on the proximal p...
research
04/07/2023

UAV Obstacle Avoidance by Human-in-the-Loop Reinforcement in Arbitrary 3D Environment

This paper focuses on the continuous control of the unmanned aerial vehi...
research
08/05/2021

Deep Reinforcement Learning for Continuous Docking Control of Autonomous Underwater Vehicles: A Benchmarking Study

Docking control of an autonomous underwater vehicle (AUV) is a task that...
research
06/01/2018

Multi-vehicle Flocking Control with Deep Deterministic Policy Gradient Method

Flocking control has been studied extensively along with the wide applic...
research
08/19/2019

Computational Flight Control: A Domain-Knowledge-Aided Deep Reinforcement Learning Approach

This papers aims to examine the potential of using the emerging deep rei...
research
08/02/2021

Three-Dimensional Trajectory Design for Multi-User MISO UAV Communications: A Deep Reinforcement Learning Approach

In this paper, we investigate a multi-user downlink multiple-input singl...
research
09/11/2017

Autonomous Quadrotor Landing using Deep Reinforcement Learning

Landing an unmanned aerial vehicle (UAV) on a ground marker is an open p...

Please sign up or login with your details

Forgot password? Click here to reset