arrow
返回

Multi-objective fuzzy Q-learning to solve continuous state-action problems

delete2023-01-01
delete8
PRE
AI
A
Amirhossein Asgharnia *
H
Howard M. Schwartz
M
Mohamed Atia
DOI:10.1016/j.neucom.2022.10.035delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Many real world problems are multi-objective. Thus, the need for multi-objective learning and optimiza-tion algorithms is inevitable. Although the multi-objective optimization algorithms are well-studied, the multi-objective learning algorithms have attracted less attention. In this paper, a fuzzy multi-objective reinforcement learning algorithm is proposed, and we refer to it as the multi-objective fuzzy Q-learning (MOFQL) algorithm. The algorithm is implemented to solve a bi-objective reach-avoid game. The majority of the multi-objective reinforcement algorithms proposed address solving problems in the discrete state-action domain. However, the MOFQL algorithm can also handle problems in a contin-uous state-action domain. A fuzzy inference system (FIS) is implemented to estimate the value function for the bi-objective problem. We used a temporal difference (TD) approach to update the fuzzy rules. The proposed method isa multi-policy multi-objective algorithm and can find the non-convex regions of the Pareto front.(c) 2022 Elsevier B.V. All rights reserved.
Keyword:
Reinforcement learning
Differential games
Q-learning
Multi-objective reinforcement learning

期刊

Neurocomputing 封面图
Neurocomputing
IF:
6.5
论文数:
2.5W
被引数:
6.5W

机构

C
carleton university
学者数:
7.5K
论文数: 8.3K
被引数: 5
引用论文

引用论文

A temporal difference method for multi-objective reinforcement learning
err2017-11-01
err22
errOAAI
errRuiz-Montiel, Manuela; Mandow, Lawrence; Perez-de-la-Cruz, Jose-Luis
err分享
err收藏
Non‐invasive methods for the assessment of brown adipose tissue in humans
err2018-01-15
err0
errOAAI
errMaria Chondronikola; Scott C. Beeman; Richard L. Wahl
err分享
err收藏
err
IF0
err
err0
errOAAI
err
err分享
err收藏
学者 查看更多内容