arrow
返回

Learning optimal decisions with confidence

delete2019-11-15
delete42
delete
OA
AI
J
Jan Drugowitsch *
A
André G. Mendonça
Z
Zachary F. Mainen
A
Alexandre Pouget
DOI:10.1073/pnas.1906787116delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Diffusion decision models (DDMs) are immensely successful models for decision making under uncertainty and time pressure. In the context of perceptual decision making, these models typically start with two input units, organized in a neuron-antineuron pair. In contrast, in the brain, sensory inputs are encoded through the activity of large neuronal populations. Moreover, while DDMs are wired by hand, the nervous system must learn the weights of the network through trial and error. There is currently no normative theory of learning in DDMs and therefore no theory of how decision makers could learn to make optimal decisions in this context. Here, we derive such a rule for learning a near-optimal linear combination of DDM inputs based on trial-by-trial feedback. The rule is Bayesian in the sense that it learns not only the mean of the weights but also the uncertainty around this mean in the form of a covariance matrix. In this rule, the rate of learning is proportional (respectively, inversely proportional) to confidence for incorrect (respectively, correct) decisions. Furthermore, we show that, in volatile environments, the rule predicts a bias toward repeating the same choice after correct decisions, with a bias strength that is modulated by the previous choice's difficulty. Finally, we extend our learning rule to cases for which one of the choices is more likely a priori, which provides insights into how such biases modulate the mechanisms leading to optimal decisions in diffusion models.
Keyword:
decision making
diffusion models
optimality
confidence
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

P
Proceedings of the National Academy of Sciences of the United States of America
IF:
9.1
论文数:
10.8W
被引数:
73.5W

机构

H
Harvard University
学者数:
26.5W
论文数: 22.0W
被引数: 28.7W
F
fundacao champalimaud
学者数:
1.2K
论文数: 910
被引数: 4
H
Harvard Medical School
学者数:
6.5W
论文数: 4.8W
被引数: 91
学者 查看更多机构
引用论文

引用论文

Bayesian theories of conditioning in a changing world变化世界中的贝叶斯条件理论
err2006-07-01
err350
PREAI
errCourville, Aaron C.; Daw, Nathaniel D.; Touretzky, David S.
err分享
err收藏
Low Loss Plastic Terahertz Photonic Band-Gap Fibres
err2008-10-30
err0
PREAI
errGeng You-Fu; Tan Xiao-Ling; Zhong Kai; Wang Peng; Yao Jian-Quan
err分享
err收藏
Anti-Platelet Factor 4/Heparin Antibody Formation Occurs Endogenously and at Unexpected High Frequency in Polycythemia Vera
err2017-01-01
err0
errOAAI
errSara C. Meyer; Eva Steinmann; Thomas Lehmann; Patricia Muesser; Jakob R. Passweg; Radek C. Skoda; Dimitrios A. Tsakiris
err分享
err收藏
Tuning the speed-accuracy trade-off to maximize reward rate in multisensory decision-making
err2015-06-19
err55
errOAAI
errDrugowitsch, Jan; DeAngelis, Gregory C.; Angelaki, Dora E.; Pouget, Alexandre
err分享
err收藏
The overlap model: A model of letter position coding
err2008-01-01
err361
errOAAI
errGomez, Pablo; Ratcliff, Roger; Perea, Manuel
err分享
err收藏
Neural correlations, population coding and computation
err2006-05-01
err1.4K
PREAI
errAverbeck, BB; Latham, PE; Pouget, A
err分享
err收藏
err分享
err收藏
Low Complexity Detection for Quadrature Spatial Modulation Systems
err2017-03-04
err0
PREAI
errJun Li; Xueqin Jiang; Yier Yan; Wenjun Yu; Sangseob Song; Moon Ho Lee
err分享
err收藏
Optimal policy for value-based decision-making
err2016-08-18
err129
errOAAI
errTajima, Satohiro; Drugowitsch, Jan; Pouget, Alexandre
err分享
err收藏
学者 查看更多内容