arrow
返回

Learning a decision maker's utility function from (possibly) inconsistent behavior

delete2004-12-01
delete35
PRE
AI
T
Thomas D. Nielsen
F
Finn V. Jensen
DOI:10.1016/j.artint.2004.08.003delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
When modeling a decision problem using the influence diagram framework, the quantitative part rests on two principal components: probabilities for representing the decision maker's uncertainty about the domain and utilities for representing preferences. Over the last decade, several methods have been developed for learning the probabilities from a database. However, methods for learning the utilities have only received limited attention in the computer science community. A promising approach for learning a decision maker's utility function is to take outset in the decision maker's observed behavioral patterns, and then find a utility function which (together with a domain model) can explain this behavior. That is, it is assumed that decision maker's preferences are reflected in the behavior. Standard learning algorithms also assume that the decision maker is behavioral consistent, i.e., given a model of the decision problem, there exists a utility function which can account for all the observed behavior. Unfortunately, this assumption is rarely valid in real-world decision problems, and in these situations existing learning methods may only identify a trivial utility function. In this paper we relax this consistency assumption, and propose two algorithms for learning a decision maker's utility function from possibly inconsistent behavior; inconsistent behavior is interpreted as random deviations from an underlying (true) utility function. The main difference between the two algorithms is that the first facilitates a form of batch learning whereas the second focuses on adaptation and is particularly well-suited for scenarios where the DM's preferences change over time. Empirical results demonstrate the tractability of the algorithms, and they also show that the algorithms converge toward the true utility function for even very small sets of observations. (C) 2004 Elsevier B.V. All rights reserved.
Keyword:
influence diagram
learning utility functions
inconsistent behavior
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Artificial Intelligence Review 封面图
Artificial Intelligence Review
IF:
13.9
论文数:
6.1K
被引数:
1.9W

机构

暂无机构信息
引用论文

引用论文

Positive interactions lead to lasting positive memories in horses, Equus caballus
err2010-04-01
err0
PREAI
errCarol Sankey; Marie-Annick Richard-Yris; Hélène Leroy; Séverine Henry; Martine Hausberger
err分享
err收藏
Neuronal correlates of asocial behavior in a BTBR T+Itpr3tf/J mouse model of autism
err2015-08-06
err0
errOAAI
errKsenia Meyza; Tomasz Nikolaev; Kacper Kondrakiewicz; D. Caroline Blanchard; Robert J. Blanchard; Ewelina Knapska
err分享
err收藏
The career decisions of young men
err1997-06-01
err687
errOAAI
errKeane, MP; Wolpin, KI
err分享
err收藏
Evolutionary Models of Rotating Stars
err1991-01-01
err0
PREAI
errSabatino Sofia; Marc Pinsonneault; Constantine P. Deliyannis
err分享
err收藏
err分享
err收藏
err分享
err收藏
没有更多内容