arrow
返回

A Novel Tree-Based Method for Interpretable Reinforcement Learning

delete2024-09-09
delete0
PRE
AI
Y
Yifan Li *
S
Shuhan Qi
王
王晅 (Xuan Wang)
J
Jiajia Zhang
L
Lei Cui
DOI:10.1145/3695464delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Deep reinforcement learning (DRL) has garnered remarkable success across various domains, propelled by advancements in deep learning (DL) technologies. However, the opacity of DL presents significant challenges, limiting the application of DRL in critical systems. In response, decision tree (DT)-based methods, known for their transparent decision-making mechanisms, have shown promise in making interpretable policies for decision-making problems. Existing methods often employ differential DTs to model RL policies and discretize them to conventional DTs for higher interpretability. Yet, this method leads to discrepancies between the trained differential DTs and the discretized DTs. To address this issue, we introduce Generative Consistent Trees (GCTs), a novel solution that circumvents the information loss typically associated with the argmax operation in prior research. By implementing a reparameterization technique to approximate the categorical distribution, GCTs ensure the consistencies between trained GCTs and discretized counterparts. Moreover, we have developed an imitation learning-based framework for interpretable reinforcement learning. This framework is designed to train GCTs by efficiently mimicking expert policies. Our extensive experiments across multiple environments have validated the effectiveness of this approach, highlighting the potential of GCTs in enhancing the interpretability and applicability of DRL.
Keyword:
Interpretable reinforcement learning
decision tree
decision-making

期刊

ACM Transactions on Knowledge Discovery from Data 封面图
ACM Transactions on Knowledge Discovery from Data
IF:
4.8
论文数:
1.3K
被引数:
4.4K

机构

H
harbin institute of technology
学者数:
8.0W
论文数: 6.6W
被引数: 66
引用论文

引用论文

Ethology and psychotherapy
err1994-09-01
err0
PREAI
errTyge Schelde; Mogens Hertz
err分享
err收藏
Trustworthy AI: A Computational Perspective值得信赖的人工智能: 计算视角
err2022-11-09
err51
errOAAI
errLiu, Haochen; Wang, Yiqi; Fan, Wenqi; Liu, Xiaorui; Li, Yaxin; Jain, Shaili; Liu, Yunhao; Jain, Anil; Tang, Jiliang
err分享
err收藏
Adversarial Attacks for Black-Box Recommender Systems via Copying Transferable Cross-Domain User Profiles
err2023-12-01
err3
PREAI
errFan, Wenqi; Zhao, Xiangyu; Li, Qing; Derr, Tyler; Ma, Yao; Liu, Hui; Wang, Jianping; Tang, Jiliang
err分享
err收藏
An Integrated Multi-Task Model for Fake News Detection
err2022-11-01
err85
PREAI
errLiao, Qing; Chai, Heyan; Han, Hao; Zhang, Xiang; Wang, Xuan; Xia, Wen; Ding, Ye
err分享
err收藏
err分享
err收藏
学者 查看更多内容