arrow
返回

A Cognitive Jamming Decision-Making Method Based on Heuristic Improved A2C Algorithm

delete2025-02-01
delete0
PRE
AI
C
Chudi Zhang
B
Biao Yang
王磊 封面图
王磊 (Lei Wang)
W
Wenshuai Ji
王璐璐 封面图
王璐璐 (Lulu Wang)
S
Shiyou Xu *
DOI:10.1109/TVT.2024.3470832delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Cognitive electronic warfare (CEW) has received increasing attention, and it is widely recognized that it will play a significant role. Cognitive jamming decision-making, as one of the critical technologies of CEW, dramatically impacts the global battlefield situation. In this paper, we introduce the A2C algorithm into cognitive jamming decision-making and propose a heuristic improved A2C algorithm. As a fusion algorithm of DQN and Policy Gradient, on the one hand, the Actor-Critic algorithm has the advantages of iterative updating of policies and high efficiency in complex spaces, compared to the DQN algorithm. On the other hand, compared with the Policy Gradient algorithm, it features fast convergence. But, it suffers from high variance, making convergence difficult. First, we establish a cognitive jamming decision- making model to address the above issues. Then, we develop an improved A2C algorithm by introducing a baseline and dueling networks. The baseline reduces the variance of the critic network, while dueling networks further decrease variance, enhancing the convergence of the A2C algorithm. Additionally, the improved A2C algorithm does not rely on prior information, and enhances the adaptive capability of the jammer when interacting with the target radar. We conducted numerical simulations based on the designed cognitive jamming decision-making model. The results demonstrated that compared with the four algorithms (DQN, Policy Gradient, Actor-Critic, A2C), the convergence speed of the improved A2C algorithm is improved by 50%, 57.8%, 34.5%, and 13.64%, respectively, and verified the excellent performance of the improved A2C. Finally, we introduce the heuristic reward function and propose the heuristic improved A2C algorithm. Compared with the improved A2C algorithm, the convergence speed of this algorithm is improved by 31.58%. The simulation results prove that the algorithm can greatly improve our advantage in electronic countermeasures.
Keyword:
Jamming
Radar
Q-learning
Decision making
Heuristic algorithms
Convergence
Artificial intelligence
Adaptation models
Airborne radar
Cognitive radar
Actor-Critic
and heuristic improved A2C
Cognitive electronic warfare
jamming decision-making
reinforcement learning

期刊

IEEE Transactions on Vehicular Technology 封面图
IEEE Transactions on Vehicular Technology
IF:
7.1
论文数:
1.8W
被引数:
6.6W

机构

S
Sun Yat Sen University
学者数:
9.9W
论文数: 7.2W
被引数: 95
引用论文

引用论文

Creating the Capacity to Screen Deaf Women for Perinatal Depression: A Pilot Study
err2021-01-01
err0
errOAAI
errMelissa L. Anderson; Kelly S. Wolf Craig; Sheri Hostovsky; Maureen Bligh; Emily Bramande; Kristin Walker; Kathleen Biebel; Nancy Byatt
err分享
err收藏
err分享
err收藏
An Overview of Cognitive Radar: Past, Present, and Future
err2019-12-01
err111
PREAI
errGurbuz, Sevgi Zubeyde; Griffiths, Hugh D.; Charlish, Alexander; Rangaswamy, Muralidhar; Greco, Maria Sabrina; Bell, Kristine
err分享
err收藏
Factors influencing women's satisfaction with surgical abortion
err2016-02-01
err0
errOAAI
errCandice Tilles; Ashleigh Denny; Catherine Cansino; Mitchell D. Creinin
err分享
err收藏
err分享
err收藏
Perioperative outcomes and disparities in utilization of sentinel lymph node biopsy in minimally invasive staging of endometrial cancer
err2020-12-01
err0
PREAI
errBenjamin B. Albright; Dimitrios Nasioudis; Maureen E. Byrne; Nawar A. Latif; Emily M. Ko; Ashley F. Haggerty
err分享
err收藏
学者 查看更多内容