arrow
Return

Nash Q-Network for Multi-agent Cybersecurity Simulation

delete2026-01-01
delete0
PRE
AI
Q
Qintong Xie *
E
Edward Koh
X
Xavier F. Cadet
C
Chin, Peter
DOI:10.1007/978-3-032-08067-7_3delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Cybersecurity defense involves interactions between adversarial parties (namely defenders and hackers), making multi-agent reinforcement learning (MARL) an ideal approach for modeling and learning strategies for these scenarios. This paper addresses the challenge of simultaneous multi-agent training in complex environments and introduces a Nash Q-Network that enables learning in a partial observation environment. Facilitates learning in partially observed settings. We demonstrate the successful implementation of this algorithm in a notable complex cyber defense simulation treated as a two-player zero-sum Markov game setting. We propose the Nash Q-Network, which aims to learn Nash-optimal strategies that translate to robust defenses in cybersecurity settings. Our approach incorporates aspects of proximal policy optimization (PPO), deep Q-network (DQN), and the Nash-Q algorithm, addressing common challenges like non-stationarity and instability in multi-agent learning. The training process employs distributed data collection and carefully designed neural architectures for both agents and critics.
Keywords:
Cybersecurity
Game Theory
Reinforcement Learning

Journal

G
GAME THEORY AND AI FOR SECURITY, GAMESEC 2025, PT II
IF:
0
Papers:
16
Citations:
0

Organization

D
dartmouth college
Scholars:
1.8K
Papers: 852
Citations: 1