arrow
返回

Boosting Temporal Binary Coding for Large-Scale Video Search

delete2021-01-01
delete10
PRE
AI
Y
Yan Wu
X
Xianglong Liu *
H
Haotong Qin
胡
胡胜 (Sheng Hu)
Y
Yuqing Ma
王
王萌 (Meng Wang)
DOI:10.1109/TMM.2020.2978593delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
In recent years, there has been an explosive increase in the amount of existing visual data. Hashing techniques have been successfully applied to deal with the large-scale nearest neighbor search problem among data on this massive scale. However, existing hashing methods usually learn a single hash code for each data point, and only by taking the content correlations among them into account. In practice, however, when handling complex visual data such as video, strong temporal relations exist among the successive frames. Moreover, if the preferred performance for large-scale video search is to be delivered, multiple hash codes are required for each data point in order to build multiple hash table indices. To address these problems, in this paper, we first study the multi-table learning problem for video search and attempt to learn binary codes by capturing the intrinsic video similarities from both the visual and the temporal aspects. By regarding the search over multiple tables as an ensemble prediction, the whole multi-table learning problem can be solved in a boosting learning manner to complementarily cover the nearest neighbors. For each table, a temporal binary coding solution is devised that thinks over the intrinsic relations among the visual content and the temporal consistency among the successive frames simultaneously. More specifically, we approximate the intrinsic visual similarities using a low-rank matrix based on sparse, non-negative feature expression. Furthermore, to essentially preserve the temporal consistency, we introduce a subspace rotation to model the variation among the successive frames. Under the boosting learning framework, the binary codes, hash functions and temporal variation of each table can be efficiently and jointly optimized. Extensive experiments on three large video datasets demonstrate that the proposed approach significantly outperforms a number of state-of-the-art hashing methods.
Keyword:
Visualization
Binary codes
Boosting
Encoding
Semantics
Indexing
Feature extraction
Large-scale video search
binary code learning
locality-sensitive hashing
temporal consistency
multi-table indexing
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Transactions on Multimedia 封面图
IEEE Transactions on Multimedia
IF:
9.7
论文数:
4.5K
被引数:
2.4W

机构

H
hefei university of technology
学者数:
2.5W
论文数: 1.7W
被引数: 35
B
Beihang University
学者数:
5.2W
论文数: 4.1W
被引数: 37
引用论文

引用论文

The 1965 Eruption of Taal Volcano
err1966-02-25
err0
PREAI
errJames G. Moore; Kazuaki Nakamura; Arturo Alcaraz
err分享
err收藏
Deer Response to a Drive Census Determined by Radio Tracking
err1965-02-01
err0
PREAI
errJohn R. Tester; Keith L. Heezen
err分享
err收藏
err分享
err收藏
Semantic Neighbor Graph Hashing for Multimodal Retrieval
err2018-03-01
err35
PREAI
errJin, Lu; Li, Kai; Hu, Hao; Qi, Guo-Jun; Tang, Jinhui
err分享
err收藏
err分享
err收藏
err分享
err收藏
Heterogeneous Hashing Network for Face Retrieval Across Image and Video Domains
err2019-03-01
err24
PREAI
errJing, Chenchen; Dong, Zhen; Pei, Mingtao; Jia, Yunde
err分享
err收藏
Propagation of acoustic waves in nematic elastomers
err2002-11-20
err0
PREAI
errE. M. Terentjev; I. V. Kamotski; D. D. Zakharov; L. J. Fradkin
err分享
err收藏
学者 查看更多内容