Skip to main navigation Skip to search Skip to main content

A UAV Path Planning Method in Three-Dimensional Urban Airspace based on Safe Reinforcement Learning

  • Yan Li*
  • , Xuejun Zhang
  • , Yuanjun Zhu
  • , Ziang Gao
  • *Corresponding author for this work
  • Beihang University

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

Abstract

Under the demand of urban terminal "Last Mile Delivery"scenario, finding a safe and efficient UAV path planning method is a crucial issue of current research. Nowadays, reinforcement learning is widely used in UAV path planning, but it is difficult to ensure the safety of the learning or execution phases due to the lack of hard constraints. Aiming at the constraints above, this paper studies how to combine safety properties with RL algorithm to find a safe path and proposes a safe reinforcement learning method called Shield-DDPG for UAV path planning. In the method, a protection mechanism Shield is mainly introduced to prevent the algorithm from outputting unsafe actions. Further, the state space, action space, and reward function are specifically improved for efficiency and safety. Then we compare the Shield-DDPG algorithm with the DDPG and RRT algorithm in some different scenarios, and the results show that the proposed algorithm has a better performance. With the proposed path planning method, UAV can learn well to efficiently and safely reach the destination via calling the trained policy. This research is of great importance to UAV operations and practical applications in complex urban airspace.

Original languageEnglish
Title of host publicationDASC 2023 - Digital Avionics Systems Conference, Proceedings
PublisherInstitute of Electrical and Electronics Engineers Inc.
ISBN (Electronic)9798350333572
DOIs
StatePublished - 2023
Event42nd IEEE/AIAA Digital Avionics Systems Conference, DASC 2023 - Barcelona, Spain
Duration: 1 Oct 20235 Oct 2023

Publication series

NameAIAA/IEEE Digital Avionics Systems Conference - Proceedings
ISSN (Print)2155-7195
ISSN (Electronic)2155-7209

Conference

Conference42nd IEEE/AIAA Digital Avionics Systems Conference, DASC 2023
Country/TerritorySpain
CityBarcelona
Period1/10/235/10/23

Keywords

  • Deep Deterministic Policy Gradient (DDPG)
  • Path Planning
  • Safe Reinforcement Learning (SRL)
  • Shield
  • Unmanned Aerial Vehicle (UAV)

Fingerprint

Dive into the research topics of 'A UAV Path Planning Method in Three-Dimensional Urban Airspace based on Safe Reinforcement Learning'. Together they form a unique fingerprint.

Cite this