arrow
返回

Deep reinforcement learning as multiobjective optimization benchmarks: Problem formulation and performance assessment

delete2024-10-01
delete0
PRE
AI
O
Oladayo S. Ajani
D
Dzeuban Fenyom Ivan
D
Daison Darlan
P
Ponnuthurai Nagaratnam Suganthan
K
Kaizhou Gao
R
Rammohan Mallipeddi *
DOI:10.1016/j.swevo.2024.101692delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
The successful deployment of Deep learning in several challenging tasks has been translated into complex control problems from different domains through Deep Reinforcement Learning (DRL). Although DRL has been extensively formulated and solved as single-objective problems, nearly all real-world RL problems often feature two or more conflicting objectives, where the goal is to obtain a high-quality and diverse set of optimal policies for different objective preferences. Consequently, the development of Multi-Objective Deep Reinforcement Learning (MODRL) algorithms has gained a lot of traction in the literature. Generally, Evolutionary Algorithms (EAs) have been demonstrated to be scalable alternatives to the classical DRL paradigms when formulated as an optimization problem. Hence it is reasonable to employ Multi-objective Evolutionary Algorithms (MOEAs) to handle MODRL tasks. However, there are several factors constraining the progress of research along this line: first, there is a lack of a general problem formulation of MODRL tasks from an optimization perspective; second, there exist several challenges in performing benchmark assessments of MOEAs for MODRL problems. To overcome these limitations: (i) we present a formulation of MODRL tasks as general multi-objective optimization problems and analyze their complex characteristics from an optimization perspective; (ii) we present an end-to-end framework, termed DRLXBench, to generate MODRL benchmark test problems for seamless running of MOEAs (iii) we propose a test suite comprising of 12 MODRL problems with different characteristics such as many-objectives, degenerated Pareto fronts, concave and convex optimization problems, etc. (iv) Finally, we present and discuss baseline results on the proposed test problems using seven representative MOEAs.
Keyword:
Evolutionary multi-objective optimization
Multi-objective reinforcement learning
Neuroevolution

期刊

Swarm and Evolutionary Computation 封面图
Swarm and Evolutionary Computation
IF:
8.5
论文数:
2.2K
被引数:
1.0W

机构

K
kyungpook national university (knu)
学者数:
1.8W
论文数: 1.8W
被引数: 14
Q
Qatar University
学者数:
8.9K
论文数: 9.0K
被引数: 16
引用论文

引用论文

A Conceptual Framework for Servitization in Industry 4.0: Distilling Directions for Future Research
err2020-01-01
err0
errOAAI
errCaroline Ennis; Nicholas Barnett; Sergio De Cesare; Rachel Lander; Alan Pilkington
err分享
err收藏
err分享
err收藏
Indicator-based Multi-objective Evolutionary Algorithms: A Comprehensive Survey
err2020-03-20
err105
PREAI
errGuillermo Falcon-Cardona, Jesus; Coello Coello, Carlos A.
err分享
err收藏
Reinforcement learning-assisted evolutionary algorithm: A survey and research opportunities强化学习辅助的进化算法: 调查与研究机会
err2024-04-01
err27
errOAAI
errSong, Yanjie; Wu, Yutong; Guo, Yangyang; Yan, Ran; Suganthan, Ponnuthurai Nagaratnam; Zhang, Yue; Pedrycz, Witold; Das, Swagatam; Mallipeddi, Rammohan; Ajani, Oladayo Solomon; Feng, Qiang
err分享
err收藏
学者 查看更多内容