arrow
Return

Multi-objective reinforcement learning for fed-batch fermentation process control

delete2022-07-01
delete10
PRE
AI
D
Dazi Li *
F
Fuqiang Zhu
X
Xiao Wang
靳其兵 (Qibing Jin)
DOI:10.1016/j.jprocont.2022.05.003delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Many real-world control problems involve conflicting objectives. For different objectives, it is necessary to obtain Pareto optimal solution sets for each one. Over recent years, multi-objective reinforcement learning (MORL) has been extensively studied to solve this problem. However, the multi-objective optimization problem of complex continuous control processes still requires exploration. Soft proximal policy optimization algorithms have also been proposed to be combined with a hybrid weightgeneration method for application to find the Pareto front approximation of the fed-batch fermentation process. This algorithm intends to initially find a single policy for the multi-objective reinforcement learning problem. A hybrid weight-generation method is then used to change the weights between different objectives to find a set of Pareto optimal solutions. In addition, we analyzed the mechanism of the fed-batch process, established the kinetic model, and designed the experimental environment based on the OpenAI Gym library. Experimental results showed that the proposed algorithm is effective and efficient in approaching the Pareto front of the fed-batch fermentation problem. (c) 2022 Elsevier Ltd. All rights reserved.
Keywords:
Multi-objective
Reinforcement learning
Fed-batch fermentation
Process control

Journal

Journal of Process Control cover
Journal of Process Control
IF:
3.9
Papers:
3.5K
Citations:
7.3K

Organization

B
Beijing University of Chemical Technology
Scholars:
3.1W
Papers: 2.2W
Citations: 4.5W
Cited Papers

Cited Papers

Aqueous phase epoxidation of 1-butene catalyzed by suspension of Au/TiO2 +TS-1
err2009-12-01
err0
errOAAI
errJian Jiang; Harold H. Kung; Mayfair C. Kung; Jiantai Ma
errShare
errSave
Performance assessment of multiobjective optimizers: An analysis and review
err2003-04-01
err3.1K
errOAAI
errZitzler, E; Thiele, L; Laumanns, M; Fonseca, CM; da Fonseca, VG
errShare
errSave
errShare
errSave
NON-SEASONALITY OF SICKLE-CELL CRISIS
err1973-09-01
err0
PREAI
errRuthAndrea Seeler
errShare
errSave
Asenapine reduces anxiety-related behaviours in rat conditioned fear stress model
err2016-04-21
err0
PREAI
errMasayo Ohyama; Maho Kondo; Miki Yamauchi; Taiichiro Imanishi; Tsukasa Koyama
errShare
errSave
errShare
errSave
researcher View more