arrow
Return

Combining knowledge with data for efficient and generalizable visual learning

delete2019-06-01
delete4
PRE
AI
Q
Qiang Ji *
DOI:10.1016/j.patrec.2017.11.013delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Substantial progress has been made in the past decades in computer vision, in particular as a result of the application of deep learning methods. Despite these rapid developments, there still exist a significant gap between computer vision and human vision. One factor contributing to this gap is the data-driven and purely bottom-up nature of the existing visual learning methods and their inability to use prior knowledge. The data-driven and bottom-up approaches often cannot generalize well beyond the data that is used to train them. Parallel to data, there is usually prior knowledge in many domains that governs the target object, its context, and the computer vision tasks. Such knowledge, if utilized properly, can not only improve visual recognition performance but also reduce our dependence on data. To this goal, we propose to identify the related prior knowledge from different sources and to systematically encode them into visual learning tasks though joint bottom-up and top-down inference. Specifically, we first identify four types of prior knowledge, including permanent theoretical knowledge, circumstantial knowledge, subjective experiential knowledge, and data knowledge. We then demonstrate how permanent theoretical knowledge and circumstantial knowledge can be identified for different vision tasks and introduce methods to systematically represent and integrate them with the image data. Experiments on benchmark datasets show that by employing related prior knowledge, we can produce vision algorithms that are more data efficient, robust and generalizable and that are less dependent on training data. (C) 2017 Elsevier B.V. All rights reserved.
Keywords:
Computer vision
Machine learning
Object recognition
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Pattern Recognition Letters cover
Pattern Recognition Letters
IF:
3.3
Papers:
8.0K
Citations:
1.6W

Organization

R
rensselaer polytechnic institute
Scholars:
7.0K
Papers: 6.5K
Citations: 6
Cited Papers

Cited Papers

A new learning paradigm: Learning using privileged information
err2009-07-01
err564
PREAI
errVapnik, Vladimir; Vashist, Akshay
errShare
errSave
Structure of the human spermidine/spermine N1-acetyltransferase gene
err1992-09-01
err0
errOAAI
errLei Xiao; Celano Paul; Amy R. Mank; Constance Griffin; Ethylin Wang Jabs; Anita L. Hawkins; Robert A. Casero
errShare
errSave
The (European) Derisking State
err
IF0
err2023-05-17
err0
errOAAI
errDaniela Gabor
errShare
errSave
Progress of Organic/Inorganic Luminescent Materials for Optical Wireless Communication Systems
err2023-06-07
err0
errOAAI
errJavier Martínez; Igor Osorio-Roman; Andrés F. Gualdrón-Reyes
errShare
errSave
Dynamics of facial expression extracted automatically from video
err2006-06-01
err225
PREAI
errLittlewort, Gwen; Bartlett, Marian Stewart; Fasel, Ian; Susskind, Joshua; Movellan, Javier
errShare
errSave
Data-Free Prior Model for Facial Action Unit Recognition
err2013-04-01
err56
PREAI
errLi, Yongqiang; Chen, Jixu; Zhao, Yongping; Ji, Qiang
errShare
errSave
no more