返回
Correcting flawed expert knowledge through reinforcement learning
DOI:10.1016/j.eswa.2015.04.015.png)
摘要
En 中文
Subject matter experts can sometimes provide incorrect and/or incomplete knowledge in the process of building intelligent systems. Other times, the expert articulates correct knowledge only to be misinterpreted by the knowledge engineer. In yet other cases, changes in the domain can lead to outdated knowledge in the system. This paper describes a technique that improves a flawed tactical agent by revising its knowledge through practice in a simulated version of its operational environment. This form of theory revision repairs agents originally built through interaction with subject matter experts. It is advantageous because such systems can now cease to be completely dependent on human expertise to provide correct and complete domain knowledge. After an agent has been built in consultation with experts, and before it is allowed to become operational, our method permits its improvement by subjecting it to several practice sessions in a simulation of its mission environment. Our method uses reinforcement learning to correct such errors and fill in gaps in the knowledge of a context-based tactical agent. The method was implemented and evaluated by comparing the performance of an agent improved by our method, to the original hand-built agent whose knowledge was purposely seeded with known errors and/or gaps. The results show that the improved agent did in fact correct the seeded errors and did gain the missing knowledge to permit it to perform better than the original, flawed agent. (C) 2015 Elsevier Ltd. All rights reserved.
Keyword:
Knowledge acquisition
Context-based reasoning
Reinforcement learning
Theory revision
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
7.5
论文数:
3.0W
被引数:
10.2W
机构
引用论文
Learning in context: enhancing machine learning with context-based reasoning在上下文中学习: 通过基于上下文的推理增强机器学习
APPLIED INTELLIGENCE
IF3.5
A Dynamic-Bayesian Network framework for modeling and evaluating learning from observation用于建模和评估从观察中学习的动态贝叶斯网络框架


