arrow
Return

Multi-Agent Reinforcement Learning for a Random Access Game

delete2022-08-01
delete3
PRE
AI
D
Dongwoo Lee
Y
Yu Zhao
J
Jun-Bae Seo *
J
Joohyun Lee *
DOI:10.1109/TVT.2022.3176722delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
This work investigates a random access (RA) game for a time-slotted RA system, where N players choose a set of slots of a frame and each frame consists of M multiple time slots. We obtain the pure strategy Nash equilibria (PNEs) of this RA game, where slots are fully utilized as in the centralized scheduling. As an algorithm to realize a PNE (Pure strategy Nash Equilibrium), we propose an Exponential-weight algorithm for Exploration and Exploitation (EXP3)-based multi-agent (MA) learning algorithm, which has the computational complexity of O(N (NmaxT)-T-2). EXP3 is a bandit algorithm designed to find an optimal strategy in a multi-armed bandit (MAB) problem that users do not know the expected payoff of each strategy. Our simulation results show that the proposed algorithm can achieve PNEs. Moreover, it can adapt to time-varying environments, where the number of players varies over time.
Keywords:
Multi-armed bandit
nash equilibrium
non-cooperative game
random access

Journal

IEEE Transactions on Vehicular Technology cover
IEEE Transactions on Vehicular Technology
IF:
7.1
Papers:
1.8W
Citations:
6.6W

Organization

H
hanyang university
Scholars:
2.8W
Papers: 2.7W
Citations: 36
G
Gyeongsang National University
Scholars:
10.0K
Papers: 8.8K
Citations: 13