摘要
In this paper, the optimal output regulation problem is considered for the model-free 2-degree-of-freedom (2-DOF) helicopter. A multistep Q-learning (MsQL) method is developed with multistep policy evaluation. First, by introducing the Q-function, the optimal output regulation problem is converted to finding the optimal Q-function. Therefore, the MsQL algorithm is proposed and its convergence theory is established by showing that it generates a nonincreasing Q-function sequence that converges to the optimal Q-function. In the MsQL, the step-size of multistep policy evaluation can be different at each iteration and an adaptive tuning rule is proposed. The MsQL learns the optimal Q-function by using real system data rather than using a system model. Finally, the developed MsQL method is employed to solve the optimal output regulation problem of the model-free 2-DOF helicopter, and its effectiveness is verified.
| 源语言 | 英语 |
|---|---|
| 文章编号 | 8106728 |
| 页(从-至) | 4953-4961 |
| 页数 | 9 |
| 期刊 | IEEE Transactions on Industrial Electronics |
| 卷 | 65 |
| 期 | 6 |
| DOI | |
| 出版状态 | 已出版 - 6月 2018 |
学术指纹
探究 'Optimal Output Regulation for Model-Free Quanser Helicopter with Multistep Q-Learning' 的科研主题。它们共同构成独一无二的学术指纹。引用此
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver