Reinforcement learning reward function in unmanned aerial vehicle control tasks

1Citations
Citations of this article
13Readers
Mendeley users who have this article in their library.

This article is free to access.

Abstract

This paper presents a new reward function that can be used for deep reinforcement learning in unmanned aerial vehicle (UAV) control and navigation problems. The reward function is based on the construction and estimation of the time of simplified trajectories to the target, which are third-order Bezier curves. This reward function can be applied unchanged to solve problems in both two-dimensional and three-dimensional virtual environments. The effectiveness of the reward function was tested in a newly developed virtual environment, namely, a simplified two-dimensional environment describing the dynamics of UAV control and flight, taking into account the forces of thrust, inertia, gravity, and aerodynamic drag. In this formulation, three tasks of UAV control and navigation were successfully solved: UAV flight to a given point in space, avoidance of interception by another UAV, and organization of interception of one UAV by another. The three most relevant modern deep reinforcement learning algorithms, Soft actor-critic, Deep Deterministic Policy Gradient, and Twin Delayed Deep Deterministic Policy Gradient were used. All three algorithms performed well, indicating the effectiveness of the selected reward function.

References Powered by Scopus

Deep reinforcement learning: A brief survey

2898Citations
N/AReaders
Get full text

Reinforcement learning for demand response: A review of algorithms and modeling techniques

579Citations
N/AReaders
Get full text

Reward is enough

316Citations
N/AReaders
Get full text

Cited by Powered by Scopus

LEVIOSA: Natural Language-Based Uncrewed Aerial Vehicle Trajectory Generation

2Citations
N/AReaders
Get full text

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Cite

CITATION STYLE

APA

Tovarnov, M. S., & Bykov, N. V. (2022). Reinforcement learning reward function in unmanned aerial vehicle control tasks. In Journal of Physics: Conference Series (Vol. 2308). Institute of Physics. https://doi.org/10.1088/1742-6596/2308/1/012004

Readers over time

‘22‘23‘2502468

Readers' Seniority

Tooltip

Researcher 3

43%

PhD / Post grad / Masters / Doc 2

29%

Professor / Associate Prof. 1

14%

Lecturer / Post doc 1

14%

Readers' Discipline

Tooltip

Computer Science 4

50%

Engineering 4

50%

Save time finding and organizing research with Mendeley

Sign up for free
0