返回
A Deep Actor-Critic Reinforcement Learning Framework for Dynamic Multichannel Access
DOI:10.1109/TCCN.2019.2952909.png)
摘要
En 中文
To make efficient use of limited spectral resources, we in this work propose a deep actor-critic reinforcement learning based framework for dynamic multichannel access. We consider both a single-user case and a scenario in which multiple users attempt to access channels simultaneously. We employ the proposed framework as a single agent in the single-user case, and extend it to a decentralized multi-agent framework in the multi-user scenario. In both cases, we develop algorithms for the actor-critic deep reinforcement learning and evaluate the proposed learning policies via experiments and numerical results. In the single-user model, in order to evaluate the performance of the proposed channel access policy and the framework's tolerance against uncertainty, we explore different channel switching patterns and different switching probabilities. In the case of multiple users, we analyze the probabilities of each user accessing channels with favorable channel conditions and the probability of collision. We also address a time-varying environment to identify the adaptive ability of the proposed framework. Additionally, we provide comparisons (in terms of both the average reward and time efficiency) between the proposed actor-critic deep reinforcement learning framework, Deep-Q network (DQN) based approach, random access, and the optimal policy when the channel dynamics are known.
Keyword:
Actor-critic algorithms
channel switching patterns
deep Q-networks
deep reinforcement learning
dynamic channel access
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
I
IF:
7
论文数:
1.6K
被引数:
5.5K
机构
引用论文
A Reinforcement Learning-Based Resource Allocation Scheme for Cloud Robotics一种基于强化学习的云机器人资源分配方案
IEEE ACCESS
IF3.6
Study of bi-directional buck-boost converter topologies for application in electrical vehicle motor drives应用于电动汽车电机驱动的双向buck-boost变换器拓扑研究
On Myopic Sensing for Multi-Channel Opportunistic Access: Structure, Optimality, and Performance关于多通道机会访问的近视感知: 结构,最优性和性能

