arrow
返回

Survey on reinforcement learning for language processing

delete2022-06-03
delete45
delete
OA
AI
V
Víctor Uc-Cetina *
N
Nicolás Navarro-Guerrero
M
Martin-Gonzalez, Anabel
W
Weber, Cornelius
S
Stefan Wermter
DOI:10.1007/s10462-022-10205-5delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
In recent years some researchers have explored the use of reinforcement learning (RL) algorithms as key components in the solution of various natural language processing (NLP) tasks. For instance, some of these algorithms leveraging deep neural learning have found their way into conversational systems. This paper reviews the state of the art of RL methods for their possible use for different problems of NLP, focusing primarily on conversational systems, mainly due to their growing relevance. We provide detailed descriptions of the problems as well as discussions of why RL is well-suited to solve them. Also, we analyze the advantages and limitations of these methods. Finally, we elaborate on promising research directions in NLP that might benefit from RL.
Keyword:
Reinforcement learning
Natural language processing
Conversational systems
Parsing
Translation
Text generation

期刊

Artificial Intelligence Review 封面图
Artificial Intelligence Review
IF:
13.9
论文数:
6.1K
被引数:
1.9W

机构

U
university of hamburg
学者数:
3.7W
论文数: 2.9W
被引数: 30
U
universidad autonoma de yucatan
学者数:
1.7K
论文数: 974
被引数: 1
引用论文

引用论文

CNS neuronal focal adhesion kinase forms clusters that co-localize with vinculin
err1996-11-15
err0
errOAAI
errGerin R. Stevens; Chi Zhang; Margaret M. Berg; Mary P. Lambert; Kirsten Barber; Isabel Cantallops; Aryeh Routtenberg; William L. Klein
err分享
err收藏
Improving interactive reinforcement learning: What makes a good teacher?
err2018-03-01
err25
errOAAI
errCruz, Francisco; Magg, Sven; Nagai, Yukie; Wermter, Stefan
err分享
err收藏
err
IF0
err
err0
PREAI
err
err分享
err收藏
err
IF0
err
err0
PREAI
err
err分享
err收藏
Multimorph Eco-Evolutionary Dynamics in Structured Populations
err2022-09-01
err0
errOAAI
errSébastien Lion; Mike Boots; Akira Sasaki
err分享
err收藏
err
IF0
err
err0
PREAI
err
err分享
err收藏
学者 查看更多内容