arrow
返回

Negotiating team formation using deep reinforcement learning

delete2020-11-01
delete15
delete
OA
AI
Y
Yoram Bachrach *
R
Richard Everett
E
Edward Hughes
A
Angeliki Lazaridou
J
Joel Z. Leibo
M
Marc Lanctot
M
Michael Johanson
W
Wojciech Marian Czarnecki
T
Thore Graepel
DOI:10.1016/j.artint.2020.103356delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
When autonomous agents interact in the same environment, they must often cooperate to achieve their goals. One way for agents to cooperate effectively is to form a team, make a binding agreement on a joint plan, and execute it. However, when agents are self-interested, the gains from team formation must be allocated appropriately to incentivize agreement. Various approaches for multi-agent negotiation have been proposed, but typically only work for particular negotiation protocols. More general methods usually require human input or domain-specific data, and so do not scale. To address this, we propose a framework for training agents to negotiate and form teams using deep reinforcement learning. Importantly, our method makes no assumptions about the specific negotiation protocol, and is instead completely experience driven. We evaluate our approach on both non-spatial and spatially extended team-formation negotiation environments, demonstrating that our agents beat hand-crafted bots and reach negotiation outcomes consistent with fair solutions predicted by cooperative game theory. Additionally, we investigate how the physical location of agents influences negotiation outcomes. (C) 2020 Elsevier B.V. All rights reserved.
Keyword:
Multi-agent systems
Team formation
Coalition formation
Reinforcement learning
Deep learning
Cooperative games
Shapley value
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Artificial Intelligence Review 封面图
Artificial Intelligence Review
IF:
13.9
论文数:
6.1K
被引数:
1.9W

机构

暂无机构信息
引用论文

引用论文

Mini‐Mental State Examination
err2002-04-30
err0
PREAI
errJoseph R. Cockrell; Marshal F. Folstein
err分享
err收藏
Computing cooperative solution concepts in coalitional skill games
err2013-11-01
err28
errOAAI
errBachrach, Yoram; Parkes, David C.; Rosenschein, Jeffrey S.
err分享
err收藏
Polymer Compression in Shear Flow
err2010-06-08
err0
PREAI
errNikko Y. Chan; Ming Chen; Xiao-Tao Hao; Trevor A. Smith; Dave E. Dunstan
err分享
err收藏
err1997-01-01
err0
PREAI
errL. P. Deselliers; D. T. M. Tan; R. B. Scott; M. E. Olson
err分享
err收藏
Three-Dimensional In Vivo Kinematics of the Subtalar Joint During Dorsi-Plantarflexion and Inversion-Eversion
err2009-05-01
err0
PREAI
errAkira Goto; Hisao Moritomo; Tomonobu Itohara; Tetsu Watanabe; Kazuomi Sugamoto
err分享
err收藏
Micromeritic and Packing Properties of Diclofenac Pellets and Effects of Some Formulation Variables
err2001-06-30
err0
PREAI
errEnriqueta C. Rodriguez; J. J. Torrado; I. Nikolakakis; S. Torrado; J. L. Lastres; S. Malamataris
err分享
err收藏
Low-Temperature PECVD SiO2 On Si And SiC
err2011-02-10
err0
PREAI
errL. Teng; W. A. Anderson
err分享
err收藏
Machine learning assisted cancer cell detection using strip waveguide Bragg gratings
err2023-08-01
err0
PREAI
errNaik Parrikar Vishwaraj; Chandrika Thondagere Nataraj; Ravi Prasad Kogravalli Jagannath; Srinivas Talabattula; Gurusiddappa R. Prashanth
err分享
err收藏
Browplasty as an adjunct to rhinoplasty
err2006-07-18
err0
PREAI
errRichard C. Webster; Terence M. Davidson; Richard C. Smith
err分享
err收藏
学者 查看更多内容