arrow
返回

Feature and Region Selection for Visual Learning

delete2016-03-01
delete13
delete
OA
AI
J
Ji Zhao *
L
Liantao Wang
R
Ricardo Cabral
F
Fernando De la Torre
DOI:10.1109/TIP.2016.2514503delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Visual learning problems, such as object classification and action recognition, are typically approached using extensions of the popular bag-of-words (BoWs) model. Despite its great success, it is unclear what visual features the BoW model is learning. Which regions in the image or video are used to discriminate among classes? Which are the most discriminative visual words? Answering these questions is fundamental for understanding existing BoW models and inspiring better models for visual recognition. To answer these questions, this paper presents a method for feature selection and region selection in the visual BoW model. This allows for an intermediate visualization of the features and regions that are important for visual learning. The main idea is to assign latent weights to the features or regions, and jointly optimize these latent variables with the parameters of a classifier (e.g., support vector machine). There are four main benefits of our approach: 1) our approach accommodates non-linear additive kernels, such as the popular chi(2) and intersection kernel; 2) our approach is able to handle both regions in images and spatio-temporal regions in videos in a unified way; 3) the feature selection problem is convex, and both problems can be solved using a scalable reduced gradient method; and 4) we point out strong connections with multiple kernel learning and multiple instance learning approaches. Experimental results in the PASCAL VOC 2007, MSR Action Dataset II and YouTube illustrate the benefits of our approach.
Keyword:
Bag-of-words
feature selection
multiple kernel learning
multiple instance learning
weakly-supervised localization
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Transactions on Image Processing 封面图
IEEE Transactions on Image Processing
IF:
13.7
论文数:
1.0W
被引数:
8.4W

机构

C
Carnegie Mellon University
学者数:
1.4W
论文数: 1.4W
被引数: 2.7W
引用论文

引用论文

Comparative analysis of occlusion methods for artificial sphincters人工括约肌闭塞方法的比较分析
err2020-04-07
err0
PREAI
errLeonardo Marziale; Gioia Lucarini; Tommaso Mazzocchi; Leonardo Ricotti; Arianna Menciassi
err分享
err收藏
Designing the next generation of cryoprotectants – From proteins to small molecules
err2018-09-12
err0
PREAI
errAnna Ampaw; Thomas A. Charlton; Jennie G. Briard; Robert N. Ben
err分享
err收藏
err分享
err收藏
DemoCut
err2013-10-08
err0
errOAAI
errPei-Yu Chi; Joyce Liu; Jason Linder; Mira Dontcheva; Wilmot Li; Bjoern Hartmann
err分享
err收藏
Browsing digital video
err2000-04-01
err0
errOAAI
errFrancis C. Li; Anoop Gupta; Elizabeth Sanocki; Li-wei He; Yong Rui
err分享
err收藏
err分享
err收藏
Collages as dynamic summaries for news video
err2002-12-01
err0
PREAI
errMichael G. Christel; Alexander G. Hauptmann; Howard D. Wactlar; Tobun D. Ng
err分享
err收藏
学者 查看更多内容