摘要
To address the critical challenges of multi-agent reinforcement learning under human interactive inputs, this paper proposes a novel robust and interpretable reinforcement learning framework. First, a robustness optimization module based on an enhanced actor-critic architecture is designed to effectively mitigate disturbances arising from operational errors and subjective biases in human inputs, thereby improving system fault tolerance while ensuring policy convergence. Second, Lyapunov stability theory is innovatively incorporated into the policy optimization process, providing rigorous mathematical proofs of stability and endowing agent behaviors with white-box interpretability. Finally, by leveraging reinforcement learning design, the approach overcomes the reliance of traditional methods on continuous incentive signals, significantly enhancing algorithmic applicability in open and dynamic environments. Extensive comparative experiments validate the superior performance of the proposed method in training efficiency, policy robustness, and interpretability, offering a reliable solution for the deployment of human-machine collaborative systems in complex scenarios.
| 源语言 | 英语 |
|---|---|
| 页(从-至) | 140-145 |
| 页数 | 6 |
| 期刊 | International Conference on Robotics and Automation Sciences, ICRAS |
| 期 | 2025 |
| DOI | |
| 出版状态 | 已出版 - 2025 |
| 活动 | 9th International Conference on Robotics and Automation Sciences, ICRAS 2025 - Osaka, 日本 期限: 27 6月 2025 → 29 6月 2025 |
学术指纹
探究 'Human Interaction Reinforcement Learning With Disturbance for Muti-Agent Systems Consensus Control' 的科研主题。它们共同构成独一无二的学术指纹。引用此
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver