arrow
返回

A temporal difference method for multi-objective reinforcement learning

delete2017-11-01
delete22
delete
OA
AI
M
Manuela Ruiz-Montiel *
J
José-Luís Pérez-de-la-Cruz
DOI:10.1016/j.neucom.2016.10.100delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
This work describes MPQ-learning, an algorithm that approximates the set of all deterministic non dominated policies in multi-objective Markov decision problems, where rewards are vectors and each component stands for an objective to maximize. MPQ-learning generalizes directly the ideas of Q-learning to the multi-objective case. It can be applied to non-convex Pareto frontiers and finds both supported and unsupported solutions. We present the results of the application of MPQ-learning to some benchmark problems. The algorithm solves successfully these problems, so showing the feasibility of this approach. We also compare MPQ-learning to a standard linearization procedure that computes only supported solutions and show that in some cases MPQ-learning can be as effective as the scalarization method. (C) 2017 Elsevier B.V. All rights reserved.
Keyword:
Reinforcement learning
Multi-objective optimization
MOMDPs
Q-leaming
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Neurocomputing 封面图
Neurocomputing
IF:
6.5
论文数:
2.5W
被引数:
6.5W

机构

U
universidad de malaga
学者数:
1.2W
论文数: 9.2K
被引数: 6
引用论文

引用论文

Deficient cerebellar long-term depression and impaired motor learning in mGluR1 mutant mice
errCell
IF0
err1994-10-01
err0
PREAI
errAtsu Alba; Masanobu Kano; Chong Chen; Mark E. Stanton; Gregory D. Fox; Karl Herrup; Theresa A. Zwingman; Susumu Tonegawa
err分享
err收藏
Empirical evaluation methods for multiobjective reinforcement learning algorithms
err2010-12-22
err162
errOAAI
errVamplew, Peter; Dazeley, Richard; Berry, Adam; Issabekov, Rustam; Dekker, Evan
err分享
err收藏
Design with shape grammars and reinforcement learning使用形状语法和强化学习进行设计
err2013-04-01
err49
errOAAI
errRuiz-Montiel, Manuela; Boned, Javier; Gavilanes, Juan; Jimenez, Eduardo; Mandow, Lawrence; Perez-de-la-Cruz, Jose-Luis
err分享
err收藏
没有更多内容