Skip to main navigation Skip to search Skip to main content

HCTD3: A Hierarchical Reinforcement Learning Approach for Mixed-Action Resource Allocation in Multi-UAV Cooperative Jamming

  • Beihang University

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

Abstract

Resource allocation for multi-UAV cooperative jamming in modern electronic warfare faces significant challenges due to high-dimensional mixed action spaces, complex constraints, and dynamic environments. To address this, this paper introduces a hierarchical reinforcement learning algorithm, HCTD3. The algorithm mitigates complexity by decomposing the task into two sub-problems: Target selection (discrete actions) and power allocation (continuous actions). It employs the Gumbel-Softmax technique for high-level discrete selection and the Twin Delayed DDPG (TD3) algorithm for low-level continuous optimization, enabling efficient learning in the mixed action space. Experimental results demonstrate that, compared to traditional optimization and single-layer reinforcement learning methods, HCTD3 achieves a 37.2% average improvement in jamming effectiveness and converges 45% faster, showcasing its superior performance.

Original languageEnglish
Title of host publication2025 5th International Conference on Wireless Communication, Networking and Internet of Things, WCNIoT 2025
PublisherInstitute of Electrical and Electronics Engineers Inc.
Pages91-94
Number of pages4
ISBN (Electronic)9798331569495
DOIs
StatePublished - 2025
Event5th International Conference on Wireless Communication, Networking and Internet of Things, WCNIoT 2025 - Sydney, Australia
Duration: 5 Nov 20257 Nov 2025

Publication series

Name2025 5th International Conference on Wireless Communication, Networking and Internet of Things, WCNIoT 2025

Conference

Conference5th International Conference on Wireless Communication, Networking and Internet of Things, WCNIoT 2025
Country/TerritoryAustralia
CitySydney
Period5/11/257/11/25

Keywords

  • Multi-UAV cooperation
  • electronic jamming
  • hierarchical reinforcement learning
  • mixed action space
  • resource allocation

Fingerprint

Dive into the research topics of 'HCTD3: A Hierarchical Reinforcement Learning Approach for Mixed-Action Resource Allocation in Multi-UAV Cooperative Jamming'. Together they form a unique fingerprint.

Cite this