arrow
返回

An Improved Reinforcement Learning Method Based on Unsupervised Learning

delete2024-01-01
delete1
delete
OA
AI
X
Xin Chang *
Y
Yanbin Li
G
Guanjie Zhang
D
Donghui Liu
C
Changjun Fu
DOI:10.1109/ACCESS.2024.3351696delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
The approach of directly combining clustering method and reinforcement learning (RL) will lead to encounter the issue where states may have different state transition processes under the same action, resulting in poor policy performance. To address this challenge with multi-dimensional continuous observation data, an improved reinforcement learning method based on unsupervised learning is proposed with a novel framework. Instead of dimensionality reduction methods, unsupervised clustering is employed to indirectly capture the underlying structure of the data. First, the proposed framework incorporates multi-dimensional information, including the current observation data, the next observation data and reward information, during the clustering process, leading to a more accurate and comprehensive low-dimensional discrete representation of the observation data while retaining preserving transition of Markov decision process. Second, by compressing the observation data into a well-defined state space, the resulting cluster labels serve as the low-dimensional discrete label-states for reinforcement learning to generate more effective and robust policies. Comparative analysis with state-of-the-art RL methods demonstrates that the improved RL methods base on framework achieves higher rewards, indicating its superior performance. Furthermore, the framework exhibits computational efficiency, as evidenced by its reasonable time complexity. This structural innovation allows for better exploration and exploitation of the transition, leading to improved policy performance in engineering applications.
Keyword:
Reinforcement learning
unsupervised learning
supervised learning
deep learning
dimensionality reduction

期刊

IEEE Access 封面图
IEEE Access
IF:
3.6
论文数:
9.8W
被引数:
29.4W

机构

S
Shijiazhuang Tiedao University
学者数:
4.2K
论文数: 2.4K
被引数: 1.7K
引用论文

引用论文

Anti-Jamming Communications Using Spectrum Waterfall: A Deep Reinforcement Learning Approach
err2018-05-01
err179
errOAAI
errLiu, Xin; Xu, Yuhua; Jia, Luliang; Wu, Qihui; Anpalagan, Alagan
err分享
err收藏
err分享
err收藏
学者 查看更多内容