arrow
Return

Multi-Perspective Cost-Sensitive Context-Aware Multi-Instance Sparse Coding and Its Application to Sensitive Video Recognition

delete2016-01-01
delete18
delete
OA
AI
胡卫明 (Weiming Hu) *
X
Xinmiao Ding
李兵 (Bing Li)
王建超 cover
王建超 (Jianchao Wang)
Y
Yan Gao
F
Fangshi Wang
S
Stephen J. Maybank
DOI:10.1109/TMM.2015.2496372delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
With the development of video-sharing websites, P2P, micro-blog, mobile WAP websites, and so on, sensitive videos can be more easily accessed. Effective sensitive video recognition is necessary for web content security. Among web sensitive videos, this paper focuses on violent and horror videos. Based on color emotion and color harmony theories, we extract visual emotional features from videos. A video is viewed as a bag and each shot in the video is represented by a key frame which is treated as an instance in the bag. Then, we combine multi-instance learning (MIL) with sparse coding to recognize violent and horror videos. The resulting MIL-based model can be updated online to adapt to changing web environments. We propose a cost-sensitive context-aware multi-instance sparse coding (MI-SC) method, in which the contextual structure of the key frames is modeled using a graph, and fusion between audio and visual features is carried out by extending the classic sparse coding into cost-sensitive sparse coding. We then propose a multi-perspective multi-instance joint sparse coding (MI-J-SC) method that handles each bag of instances from an independent perspective, a contextual perspective, and a holistic perspective. The experiments demonstrate that the features with an emotional meaning are effective for violent and horror video recognition, and our cost-sensitive context-aware MI-SC and multi-perspective MI-J-SC methods outperform the traditional MIL methods and the traditional SVM and KNN-based methods.
Keywords:
Cost-sensitive context-aware multi-instance sparse coding (MI-SC)
horror video recognition
multi-perspective multi-instance joint sparse coding (MI-J-SC)
video emotional feature extraction
violent video recognition

Journal

IEEE Transactions on Multimedia cover
IEEE Transactions on Multimedia
IF:
9.7
Papers:
4.5K
Citations:
2.4W

Organization

I
institute of automation, cas
Scholars:
2.2K
Papers: 2.1K
Citations: 2
U
university of london
Scholars:
21.5W
Papers: 19.7W
Citations: 305
C
chinese academy of sciences
Scholars:
56.5W
Papers: 44.9W
Citations: 704
researcher View more organizations