arrow
返回

A Multi-Agent Approach to Modeling Task-Oriented Dialog Policy Learning

delete2025-01-01
delete0
delete
OA
AI
S
Songfeng Liang
K
Kai Xu *
Z
Zhurong Dong *
DOI:10.1109/ACCESS.2025.3529469delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Dialogue policy is a critical research area in human-computer interaction, vital for guiding dialogue generation and improving controllability and interpretability. Multi-agent dialogue policy learning demonstrates superior learning speed and exploration capabilities, positioning it as a promising approach for developing more effective and adaptive dialogue agents. However, many studies neglect to holistically model collaboration between agents, which limits the effectiveness of policy learning. Therefore, this paper proposes a new multi-agent group collaboration mechanism for dialogue policy learning, named GMPL. Concretely, we employ an Actor-Critic network to implement the proposed model, alternately updating individual dialogue agents to optimize policy selection. In each update, we utilize the maximum action value function to determine the appropriate dialogue action, while the maximum state value function serves to guide the policy learning process. This integrated approach ensures that both decision-making and learning phases are effectively aligned, thereby enhancing the overall performance of the dialogue agents. Furthermore, we conduct a theoretical analysis of the convergence properties of the proposed model. Experiments were conducted on two distinct task-oriented dialogue datasets, revealing that the proposed multi-agent model exhibits a significantly faster learning speed and a higher dialogue success rate compared to baseline approaches.
Keyword:
Natural language generation
Reinforcement learning
Decision making
Collaboration
Pipelines
Medical diagnosis
Human computer interaction
Databases
Convergence
Multi-agent systems
Human-computer interaction
dialogue policy learning
deep reinforcement learning
multi-agent learning

期刊

IEEE Access 封面图
IEEE Access
IF:
3.6
论文数:
9.8W
被引数:
29.4W

机构

S
Shenzhen Polytechnic University
学者数:
2.9K
论文数: 2.6K
被引数: 68
S
south china university of technology
学者数:
6.8W
论文数: 5.1W
被引数: 85
引用论文

引用论文

A survey on deep reinforcement learning for audio-based applications
err2022-07-02
err31
errOAAI
errLatif, Siddique; Cuayahuitl, Heriberto; Pervez, Farrukh; Shamshad, Fahad; Ali, Hafiz Shehbaz; Cambria, Erik
err分享
err收藏
学者 查看更多内容