Skip to main navigation Skip to search Skip to main content

Airship control based on Q-Learning algorithm and neural network

  • Beihang University

Research output: Contribution to journalArticlepeer-review

Abstract

An autonomous on-line learning control strategy based on adaptive modeling mechanism was proposed aimed at system modeling and parameter identification problems resulting from dynamic model uncertainties in modern airship control. An adaptive method to establish airship control Markov decision process (MDP) model was introduced on the foundation of analyzing airship's actual motion. On-line learning was carried out by Q-Learning algorithm, and cerebellar model articulation controller (CMAC) network was brought in for generalization of action value functions to accelerate algorithm convergence speed. Simulations of this autonomous on-line learning controller and comparisons with parameters turned PID controllers in normal control tasks were presented to demonstrate Q-Learning controller's effectiveness. The results show that the controller's on-line learning processes can converge in a few hours and the airship control MDP model established by the adaptive method satisfies the need of normal control tasks. The controller designed in this paper obtains similar precision as PID controllers and performs even more intelligently.

Original languageEnglish
Pages (from-to)2431-2438
Number of pages8
JournalBeijing Hangkong Hangtian Daxue Xuebao/Journal of Beijing University of Aeronautics and Astronautics
Volume43
Issue number12
DOIs
StatePublished - 1 Dec 2017

Keywords

  • Airship
  • Cerebellar model articulation controller (CMAC)
  • Machine learning
  • Markov decision process (MDP)
  • Q-Learning

Fingerprint

Dive into the research topics of 'Airship control based on Q-Learning algorithm and neural network'. Together they form a unique fingerprint.

Cite this