arrow
返回

Combining knowledge with data for efficient and generalizable visual learning

delete2019-06-01
delete4
PRE
AI
Q
Qiang Ji *
DOI:10.1016/j.patrec.2017.11.013delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Substantial progress has been made in the past decades in computer vision, in particular as a result of the application of deep learning methods. Despite these rapid developments, there still exist a significant gap between computer vision and human vision. One factor contributing to this gap is the data-driven and purely bottom-up nature of the existing visual learning methods and their inability to use prior knowledge. The data-driven and bottom-up approaches often cannot generalize well beyond the data that is used to train them. Parallel to data, there is usually prior knowledge in many domains that governs the target object, its context, and the computer vision tasks. Such knowledge, if utilized properly, can not only improve visual recognition performance but also reduce our dependence on data. To this goal, we propose to identify the related prior knowledge from different sources and to systematically encode them into visual learning tasks though joint bottom-up and top-down inference. Specifically, we first identify four types of prior knowledge, including permanent theoretical knowledge, circumstantial knowledge, subjective experiential knowledge, and data knowledge. We then demonstrate how permanent theoretical knowledge and circumstantial knowledge can be identified for different vision tasks and introduce methods to systematically represent and integrate them with the image data. Experiments on benchmark datasets show that by employing related prior knowledge, we can produce vision algorithms that are more data efficient, robust and generalizable and that are less dependent on training data. (C) 2017 Elsevier B.V. All rights reserved.
Keyword:
Computer vision
Machine learning
Object recognition
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Pattern Recognition Letters 封面图
Pattern Recognition Letters
IF:
3.3
论文数:
8.0K
被引数:
1.6W

机构

R
rensselaer polytechnic institute
学者数:
7.0K
论文数: 6.5K
被引数: 6
引用论文

引用论文

Structure of the human spermidine/spermine N1-acetyltransferase gene
err1992-09-01
err0
errOAAI
errLei Xiao; Celano Paul; Amy R. Mank; Constance Griffin; Ethylin Wang Jabs; Anita L. Hawkins; Robert A. Casero
err分享
err收藏
The (European) Derisking State
err
IF0
err2023-05-17
err0
errOAAI
errDaniela Gabor
err分享
err收藏
Progress of Organic/Inorganic Luminescent Materials for Optical Wireless Communication Systems
err2023-06-07
err0
errOAAI
errJavier Martínez; Igor Osorio-Roman; Andrés F. Gualdrón-Reyes
err分享
err收藏
Dynamics of facial expression extracted automatically from video从视频中自动提取面部表情的动态性
err2006-06-01
err225
PREAI
errLittlewort, Gwen; Bartlett, Marian Stewart; Fasel, Ian; Susskind, Joshua; Movellan, Javier
err分享
err收藏
Data-Free Prior Model for Facial Action Unit Recognition
err2013-04-01
err56
PREAI
errLi, Yongqiang; Chen, Jixu; Zhao, Yongping; Ji, Qiang
err分享
err收藏
没有更多内容