跳到主要导航 跳到搜索 跳到主要内容

Accelerating wargaming reinforcement learning by dynamic multi-demonstrator ensemble

  • Beihang University

科研成果: 期刊稿件文章同行评审

摘要

Deep Reinforcement Learning (DRL) has become a promising technique to deal with tough wargaming decision-making problems. However, DRL suffers an inherent problem of low learning efficiency and it often requires massive cost of training steps, which may be alleviated with expert demonstrations in wargaming domains. Most learning methods with demonstrations generally treat the demonstration data from different expert demonstrators without distinction. Besides, a more appropriate and effective mechanism is highly needed to control sampling balance of expert-generated demonstration samples and agent-generated interaction ones. To tackle the two issues, this work proposes an improved approach to leverage expert demonstrations to further accelerate DRL. It innovatively extracts inherent diversity in multiple demonstrators by pre-training agents individually from multiple demonstration sources, thereby producing a strong and initial ensemble model. In addition, a novel technique to evaluate the learning importance of each demonstrator is designed to dynamically tune sampling ratios of learning data in a more adaptive and effective manner. Through the evaluation on several classic game tasks and a typical wargaming scenario, our method shows superior performance over several state-of-the-art methods and significantly raises DRL's efficiency for typical wargaming decision-making applications.

源语言英语
文章编号119534
期刊Information Sciences
648
DOI
出版状态已出版 - 11月 2023

学术指纹

探究 'Accelerating wargaming reinforcement learning by dynamic multi-demonstrator ensemble' 的科研主题。它们共同构成独一无二的学术指纹。

引用此