跳到主要导航 跳到搜索 跳到主要内容

Online Human Behavior Learning via Dynamic Regressor Extension and Mixing With Fixed-Time Convergence

  • Beihang University

科研成果: 期刊稿件文章同行评审

摘要

To improve the hybrid augmented intelligence of a human-in-the-loop (HiTL) control system, it is desirable to investigate the issue of human behavior learning (HBL), i.e., empower the machine to understand how a human expert performs a manipulation task. The human expert is commonly modeled as an optimal controller with unknown weighting matrices that depict the tradeoff between different control objectives. Therefore, the goal of this article is to determine the weighting matrices of the human objective function with fast convergence rate, which is usually pursued to achieve better efficiency and performance in practice. Accordingly, for a class of HiTL system, we propose a novel adaptive inverse optimal control (IOC) approach for online learning human behavior with a fixed-time guarantee. Our proposed method consists of two parts. In the first step, a dynamic regressor extension and mixing (DREM)-based estimation method is used for online learning of the human feedback gain with fixed-time convergence using the demonstrated system state measurement only. Then, with the estimated human feedback gain, a semidefinite programming (SDP) problem is solved to determine the weighting matrices of the human objective function. The simulation and the experiment on a steering control have validated the effectiveness and applicability of the developed approach.

源语言英语
页(从-至)1764-1772
页数9
期刊IEEE Transactions on Industrial Informatics
21
2
DOI
出版状态已出版 - 2025

指纹

探究 'Online Human Behavior Learning via Dynamic Regressor Extension and Mixing With Fixed-Time Convergence' 的科研主题。它们共同构成独一无二的指纹。

引用此