arrow
返回

Semi-supervised cross-modal common representation learning with vector-valued manifold regularization

delete2020-02-01
delete7
PRE
AI
H
Hong Zhang *
T
Ting Wang
G
Gang Dai
DOI:10.1016/j.patrec.2019.01.002delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
While cross-media data, like text, image, audio, video and 3D model, has been the main form of big data, there is a current dearth of research on cross-media retrieval. In this paper, we focus on how to learn the common representation of heterogeneous data which is a key challenge for cross-media retrieval. Most existing approaches linearly project original low-level feature into a joint feature space for isomorphic data representation. However, linear projection cannot capture most complex cross-modal correlation with high nonlinearity. In this paper, we propose a novel feature learning algorithm, which is semi-supervised cross-modal vector-valued manifold regularization (SCVM), to explore common representation of heterogeneous data. SCVM jointly explores low-level feature correlation and semantic information in a unified framework. Based on manifold regularization, we learn cross-media features from vector-valued reproducing kernel Hilbert spaces (RKHS) by kernel transformation on both labeled and unlabeled samples; moreover, we impose smoothness constraints of possible solutions to improve retrieval accuracy. Comparing with the current state-of-the-art approaches on two public datasets, comprehensive experimental results show superior performance of our SCVM. The method is more robust and stable when extended from two media types to five media types, which is very attractive in practical application. (C) 2019 Published by Elsevier B.V.
Keyword:
Cross-media retrieval
Vector-valued RKHS
Manifold regularization
Semi-supervised
Kernel method

期刊

Pattern Recognition Letters 封面图
Pattern Recognition Letters
IF:
3.3
论文数:
8.0K
被引数:
1.6W

机构

暂无机构信息
引用论文

引用论文

err分享
err收藏
An Adaptive Semisupervised Feature Analysis for Video Semantic Recognition
err2018-02-01
err279
PREAI
errLuo, Minnan; Chang, Xiaojun; Nie, Liqiang; Yang, Yi; Hauptmann, Alexander G.; Zheng, Qinghua
err分享
err收藏
err分享
err收藏
Activation of mannosyltransferase II by nonbilayer phospholipids
err2002-05-01
err0
PREAI
errJohn W. Jensen; John S. Schutzbach
err分享
err收藏
Materials Imaging and Dynamics (MID) instrument at the European X-ray Free-Electron Laser Facility
err2021-02-15
err0
errOAAI
errA. Madsen; J. Hallmann; G. Ansaldi; T. Roth; W. Lu; C. Kim; U. Boesenberg; A. Zozulya; J. Möller; R. Shayduk; M. Scholz; A. Bartmann; A. Schmidt; I. Lobato; K. Sukharnikov; M. Reiser; K. Kazarian; I. Petrov
err分享
err收藏
Cross-Modal Self-Taught Hashing for large-scale image retrieval
err2016-07-01
err27
errOAAI
errXie, Liang; Zhu, Lei; Pan, Peng; Lu, Yansheng
err分享
err收藏
学者 查看更多内容