arrow
返回

PHH: Policy-Based Hyper-Heuristic With Reinforcement Learning

delete2023-01-01
delete5
delete
OA
AI
O
Orachun Udomkasemsub *
B
Booncharoen Sirinaovakul
DOI:10.1109/ACCESS.2023.3277953delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Hyper-heuristics have a high level of generality and adaptability, allowing them to effectively solve a wide range of complex optimization problems. With reinforcement learning, hyper-heuristics can use experience and knowledge gained to tackle unforeseen problems, allowing hyper-heuristics to adapt and improve over time. Our paper proposes a framework for using policy-based reinforcement learning to improve the performance of hyper-heuristics. The framework trains hyper-heuristic agents to select the best generalized constructive low-level heuristics to solve combinatorial optimization problems. The framework evaluation was performed using three benchmarking problems: traveling salesman, capacitated vehicle routing, and bin packing problems. The results showed that the proposed framework can outperform existing meta-heuristic and hyper-heuristic-based algorithms for all large problem instances in all problem domains. The proposed framework was also evaluated by applying it to a cost optimization problem for workflow scheduling on a hybrid cloud with a deadline constraint. Eight agents were trained on medium-sized workflows with two deadlines and tested against traditional meta-heuristic and hyper-heuristic methods to solve smaller and larger workflows with unforeseen deadlines. Four workflow applications, three workflow sizes, and three deadlines were used in the evaluation. The results showed that our proposed framework provided significantly better solutions of up to 98% for benchmarking problems, and up to 22% for cost optimization in workflow scheduling. Moreover, trained with small problem instances, the framework performed well for unforeseen larger problem instances implying its generalization. The proposed framework, thus, has the potential to improve both the generality and performance of solving large combinatorial optimization problems.
Keyword:
Combinatorial optimization
hyper-heuristics
policy optimization
reinforcement learning
workflow scheduling

期刊

IEEE Access 封面图
IEEE Access
IF:
3.6
论文数:
9.8W
被引数:
29.4W

机构

K
King Mongkut's University of Technology Thonburi
学者数:
3.5K
论文数: 3.6K
被引数: 3.9K
引用论文

引用论文

err分享
err收藏
Detection of minimal residual disease following induction immunochemotherapy predicts progression free survival in mantle cell lymphoma: final results of CALGB 59909
err2011-11-18
err0
errOAAI
errH. Liu; J. L. Johnson; G. Koval; G. Malnassy; D. Sher; L. E. Damon; E. D. Hsi; D. M. Bucci; C. A. Linker; B. D. Cheson; W. Stock
err分享
err收藏
Multiobjective evolutionary algorithms: A survey of the state of the art
err2011-03-01
err1.8K
PREAI
errZhou, Aimin; Qu, Bo-Yang; Li, Hui; Zhao, Shi-Zheng; Suganthan, Ponnuthurai Nagaratnam; Zhang, Qingfu
err分享
err收藏
Integrated Crop–Livestock Systems in the Texas High Plains: Productivity and Water Use
err2014-05-01
err0
errOAAI
errCody J. Zilverberg; C. Philip Brown; Paul Green; Michael L. Galyean; Vivien G. Allen
err分享
err收藏
err分享
err收藏
A Gentle Introduction to Reinforcement Learning and its Application in Different Fields
err2020-01-01
err110
errOAAI
errNaeem, Muddasar; Rizvi, Syed Tahir Hussain; Coronato, Antonio
err分享
err收藏
学者 查看更多内容