arrow
返回

LEARNING WHILE EXPERIMENTING

delete2019-07-19
delete1
PRE
AI
E
Ettore Damiano
李浩 (Hao Li)
W
Wing Suen *
DOI:10.1093/ej/uez043delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
An agent performing risky experimentation can benefit from suspending it to learn directly about the state. 'Positive' information acquisition seeks news that would confirm the state that favours experimentation. It is used as a last-ditch effort when the agent is pessimistic about the risky arm before abandoning it. 'Negative' information acquisition seeks news that would demonstrate that experimentation is futile. It is used as an insurance strategy to avoid wasteful experimentation when the agent is still optimistic. A higher reward from risky experimentation expands the region of beliefs that the agent optimally chooses information acquisition rather than experimentation.
Keyword:
DEVELOPMENT COMPETITION
DYNAMIC ALLOCATION
MODEL
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Economic Journal 封面图
Economic Journal
IF:
3.6
论文数:
5.5K
被引数:
1.6W

机构

暂无机构信息
引用论文

引用论文

err
IF0
err
err0
PREAI
err
err分享
err收藏
Strategic experimentation with exponential bandits
err2005-01-01
err258
errOAAI
errKeller, G; Rady, S; Cripps, M
err分享
err收藏
err分享
err收藏
RACING WITH UNCERTAINTY
err1987-01-01
err251
PREAI
errHARRIS, C; VICKERS, J
err分享
err收藏
Drawing Attention to the Dangerous
err2003-06-18
err0
errOAAI
errStathis Kasderidis; John G.; Nicolas Tsapatsoulis; Dario Malchiodi
err分享
err收藏
学者 查看更多内容