arrow
返回

REINFORCEMENT LEARNING IN SUPPLY CHAINS

delete2011-11-21
delete24
PRE
AI
A
Annapurna Valluri
M
Michael North *
C
Charles M. Macal
DOI:10.1142/S0129065709002063delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Effective management of supply chains creates value and can strategically position companies. In practice, human beings have been found to be both surprisingly successful and disappointingly inept at managing supply chains. The related fields of cognitive psychology and artificial intelligence have postulated a variety of potential mechanisms to explain this behavior. One of the leading candidates is reinforcement learning. This paper applies agent-based modeling to investigate the comparative behavioral consequences of three simple reinforcement learning algorithms in a multi-stage supply chain. For the first time, our findings show that the specific algorithm that is employed can have dramatic effects on the results obtained. Reinforcement learning is found to be valuable in multi-stage supply chains with several learning agents, as independent agents can learn to coordinate their behavior. However, learning in multi-stage supply chains using these postulated approaches from cognitive psychology and artificial intelligence take extremely long time periods to achieve stability which raises questions about their ability to explain behavior in real supply chains. The fact that it takes thousands of periods for agents to learn in this simple multi-agent setting provides new evidence that real world decision makers are unlikely to be using strict reinforcement learning in practice.
Keyword:
Reinforcement learning
supply chain management
agent-based modeling
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

International Journal of Neural Systems 封面图
International Journal of Neural Systems
IF:
6.4
论文数:
1.2K
被引数:
3.3K

机构

A
Argonne National Laboratory
学者数:
1.1W
论文数: 9.2K
被引数: 3.8W
U
united states department of energy (doe)
学者数:
11.3W
论文数: 9.6W
被引数: 246
引用论文

引用论文

Complexity of D″ in the presence of slab‐debris and phase changes
err2006-03-04
err0
errOAAI
errDaoyuan Sun; Teh‐Ru Alex Song; Don Helmberger
err分享
err收藏
A reinforcement learning model for supply chain ordering management: An application to the beer game
err2008-11-01
err94
PREAI
errChaharsooghi, S. Kamal; Heydari, Jafar; Zegordi, S. Hessameddin
err分享
err收藏
High-Quality Masks Reduce Covid-19 Infections and Death in the US
err2021-12-12
err0
PREAI
errErik Rosenstrom; Julie Ivy; Maria Mayorga; Julie Swann; Buse Eylul Oruc; Pinar Keskinocak; Nathaniel Hupert
err分享
err收藏
err分享
err收藏
“We Are All Gonna Get Diabetic These Days”
err2014-05-27
err0
errOAAI
errElizabeth A. Pyatak; Daniella Florindez; Anne L. Peters; Marc J. Weigensberg
err分享
err收藏
err分享
err收藏
学者 查看更多内容