arrow
返回

Rationality-bounded adaptive learning in multi-agent dynamic games

delete2023-05-01
delete1
PRE
AI
X
Xianjia Wang
X
Xue Linzhao *
Z
Zhipeng Yang
Y
Yang Liu
DOI:10.1016/j.knosys.2023.110459delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
This paper presents an adaptive learning framework that aims to enable agents with bounded rationality to learn from dynamic games and making decisions accordingly. Within this framework, agents lack complete knowledge of future payoffs but are capable of forming beliefs based on their recognitive abilities and past experiences. Each region involved dynamically adjusts its strategy by considering the tradeoff between the current payoff and belief updated. Based on this premise, we propose the Boundedly Rational Multiagent Learning (BRML) algorithm and provide proof of its convergence under given conditions. To demonstrate the practicality of the BRML algorithm, we employ it to two critical dynamic games: repeated cooperative and non-cooperative games. Our results indicate that groups with bounded rationality can achieve coordination through the behavior of agents with high recognitive ability. Furthermore, Pareto optimality is successfully attained in the coordination problems through the application of this multiagent adaptive learning framework.(c) 2023 Elsevier B.V. All rights reserved.
Keyword:
Adaptive learning
Bounded rationality
Multiagent system
Dynamic game
Coordination problem

期刊

K
Knowledge-Based Systems
IF:
7.6
论文数:
1.3W
被引数:
4.5W

机构

W
wuhan university
学者数:
8.1W
论文数: 5.8W
被引数: 70
引用论文

引用论文

Mini‐Mental State Examination
err2002-04-30
err0
PREAI
errJoseph R. Cockrell; Marshal F. Folstein
err分享
err收藏
err分享
err收藏
err分享
err收藏
Comparative Analysis of Isochoric and Isobaric Adiabatic Compressed Air Energy Storage
err2023-03-10
err0
errOAAI
errDaniel Pottie; Bruno Cardenas; Seamus Garvey; James Rouse; Edward Hough; Audrius Bagdanavicius; Edward Barbour
err分享
err收藏
Pathological outcomes of observational learning
err2000-03-01
err449
errOAAI
errSmith, L; Sorensen, P
err分享
err收藏
学者 查看更多内容