arrow
返回

Randomized sampling for large zero-sum games

delete2013-05-01
delete12
delete
OA
AI
S
Shaunak D. Bopardikar *
B
Borri, Alessandro
J
João P. Hespanha
M
Maria Prandini
M
Maria Domenica Di Benedetto
DOI:10.1016/j.automatica.2013.01.062delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
This paper addresses the solution of large zero-sum matrix games using randomized methods. We formalize a procedure, termed as the sampled security policy (SSP) algorithm, by which a player can compute policies that, with a high confidence, are security policies against an adversary using randomized methods to explore the possible outcomes of the game. The SSP algorithm essentially consists of solving a stochastically sampled subgame that is much smaller than the original game. We also propose a randomized algorithm, termed as the sampled security value (SSV) algorithm, which computes a high-confidence security-level (i.e., worst-case outcome) for a given policy, which may or may not have been obtained using the SSP algorithm. For both the SSP and the SSV algorithms we provide results to determine how many samples are needed to guarantee a desired level of confidence. We start by providing results when the two players sample policies with the same distribution and subsequently extend these results to the case of mismatched distributions. We demonstrate the usefulness of these results in a hide-and-seek game that exhibits exponential complexity. (C) 2013 Elsevier Ltd. All rights reserved.
Keyword:
Game theory
Randomized algorithms
Zero-sum games
Optimization
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Automatica 封面图
Automatica
IF:
5.9
论文数:
1.2W
被引数:
5.2W

机构

U
University of California Santa Barbara
学者数:
1.2W
论文数: 9.6K
被引数: 3.6W
University of California System 封面图
University of California System
学者数:
37.5W
论文数: 33.7W
被引数: 6.6K
R
raytheon technologies
学者数:
612
论文数: 518
被引数: 1
C
consiglio nazionale delle ricerche (cnr)
学者数:
6.2W
论文数: 5.7W
被引数: 48
学者 查看更多机构
引用论文

引用论文

The scenario approach to robust control design
err2006-05-01
err852
errOAAI
errCalafiore, Giuseppe C.; Campi, Marco C.
err分享
err收藏
Probabilistic solutions to some NP-hard matrix problems
err2001-09-01
err52
PREAI
errVidyasagar, M; Blondel, VD
err分享
err收藏
err分享
err收藏
Notes on the Scenario Design Approach
err2009-02-01
err24
PREAI
errCampi, Marco C.; Calafiore, Giuseppe C.
err分享
err收藏
学者 查看更多内容