跳到主要导航 跳到搜索 跳到主要内容

基于 PPO 的移动平台自主导航

  • Guoyan Xu*
  • , Yiwei Xiong
  • , Bin Zhou
  • , Guanhong Chen
  • *此作品的通讯作者
  • Beihang University

科研成果: 期刊稿件文章同行评审

摘要

This paper presents an autonomous navigation method based on proximal policy optimization (PPO) algorithm for mobile platform. In this method, GNSS and LADAR are used for sensing environment information. To define the state of reinforcement learning model, an ego position evaluation method is introduced based on improved artificial potential field algorithm. After that, on the basis of PPO algorithm, a kind of action policy function is designed based on Gaussian distribution, which solves the continuity problem of the vehicle linear velocity and yaw velocity. Furthermore, the network framework and reward function of the model are also designed for navigation scenarios. In order to train the navigation model, a virtual environment based on Gazebo is built. The training results show that the ego position evaluation method obviously helps to improve the speed of model convergence. Finally, the navigation model is transplanted to a real environment, which verifies the effectiveness of the proposed method.

投稿的翻译标题Autonomous navigation based on PPO for mobile platform
源语言繁体中文
页(从-至)2138-2145
页数8
期刊Beijing Hangkong Hangtian Daxue Xuebao/Journal of Beijing University of Aeronautics and Astronautics
48
11
DOI
出版状态已出版 - 11月 2022

关键词

  • artificial potential field
  • autonomous navigation
  • mobile platform
  • proximal policy optimization algorithm
  • reinforcement learning

学术指纹

探究 '基于 PPO 的移动平台自主导航' 的科研主题。它们共同构成独一无二的学术指纹。

引用此