arrow
返回

Offline Reinforcement Learning for Wireless Network Optimization With Mixture Datasets

delete2024-10-01
delete0
delete
OA
AI
K
Kun Yang *
C
Chengshuai Shi
C
Cong Shen
J
Jing Yang
S
Shu‐ping Yeh
J
Jaroslaw J. Sydir
DOI:10.1109/TWC.2024.3395624delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
The recent development of reinforcement learning (RL) has boosted the adoption of online RL for wireless radio resource management (RRM). However, online RL algorithms require direct interactions with the environment, which may be undesirable given the potential performance loss due to the unavoidable exploration in RL. In this work, we first explore the use of offline RL algorithms in solving the RRM problem. We evaluate several state-of-the-art offline RL algorithms for a practical RRM problem that aims at maximizing a linear combination of total rates and 5-percentile rates via user scheduling. Our findings indicate that the performance of offline RL for the RRM problem is heavily contingent upon the behavior policy deployed for data collection. We propose an innovative offline RL approach utilizing heterogeneous datasets from various behavior policies. This method demonstrates that a strategic mixture of datasets enables near-optimal RL policy generation, even with suboptimal behavior policies. Additionally, we introduce two enhancements: an ensemble-based policy to augment dataset mixture training efficiency, and a novel offline-to-online strategy for seamless adaptation to new environments. Our data mixture approach achieves over 95% efficiency of an online RL agent in the absence of expert data. The ensemble algorithm notably reduces training duration by half compared to the data mixture method. Furthermore, our model, when applied with offline-to-online fine-tuning, surpasses existing benchmarks by approximately 5% in our user scheduling problem.
Keyword:
Wireless networks
Resource management
Optimization
Reinforcement learning
Scheduling
Interference
Data models
Radio resource management
offline reinforcement learning
deep reinforcement learning

期刊

IEEE Transactions on Wireless Communications 封面图
IEEE Transactions on Wireless Communications
IF:
10.7
论文数:
1.3W
被引数:
5.3W

机构

U
University of Virginia
学者数:
3.0W
论文数: 2.7W
被引数: 4.1W
P
Pennsylvania State University
学者数:
3.0W
论文数: 2.6W
被引数: 7.2W
P
pennsylvania commonwealth system of higher education (pcshe)
学者数:
12.9W
论文数: 11.7W
被引数: 177
学者 查看更多机构
引用论文

引用论文

Clinical efficacy and toxicity of doxorubicin encapsulated in glutaraldehyde-treated erythrocytes administered to dogs with lymphosarcoma
err1994-06-01
err0
PREAI
errCurt M. Matherne; William C. Satterfield; Anna Gasparini; Michela Tonetti; A. Barry Astroff; Russell D. Schmidt; Loyd D. Rowe; John R. DeLoach
err分享
err收藏
Microbial modulation of behavior and stress responses in zebrafish larvae
err2016-09-01
err0
errOAAI
errDaniel J. Davis; Elizabeth C. Bryda; Catherine H. Gillespie; Aaron C. Ericsson
err分享
err收藏
Analysis on smart material suitable for autogenous microelectronic application
err2019-08-30
err0
PREAI
errR Sitharthan; Manimaran Ponnusamy; Madurakavi Karthikeyan; D Shanmuga Sundar
err分享
err收藏
学者 查看更多内容