arrow
Return

Deep Cross-Modal Proxy Hashing

delete2022-01-01
delete16
PRE
AI
R
Rong-Cheng Tu
X
Xian-Ling Mao *
C
Chengfei Cai
W
Wei Wei
黄河燕 cover
黄河燕 (Heyan Huang)
DOI:10.1109/TKDE.2022.3187023delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Due to the high retrieval efficiency and low storage cost for cross-modal search tasks, cross-modal hashing methods have attracted considerable attention from the researchers. For the supervised cross-modal hashing methods, how to make the learned hash codes sufficiently preserve semantic information contained in the label of datapoints is the key to further enhance the retrieval performance. Hence, almost all supervised cross-modal hashing methods usually depend on defining similarities between datapoints with the label information to guide the hashing model learning fully or partly. However, the defined similarity between datapoints can only capture the label information of datapoints partially and misses abundant semantic information, which then hinders the further improvement of retrieval performance. Thus, in this paper, different from previous works, we propose a novel cross-modal hashing method without defining the similarity between datapoints, called Deep Cross-modal Proxy Hashing (DCPH). Specifically, DCPH first trains a proxy hashing network to transform each category information of a dataset into a semantic discriminative hash code, called proxy hash code. Each proxy hash code can preserve the semantic information of its corresponding category well. Next, without defining the similarity between datapoints to supervise the training process of the modality-specific hashing networks, we propose a novel margin-dynamic-softmax loss to directly utilize the proxy hashing codes as supervised information. Finally, by minimizing the novel margin-dynamic-softmax loss, the modality-specific hashing networks can be trained to generate hash codes that can simultaneously preserve the cross-modal similarity and abundant semantic information well. Extensive experiments on three benchmark datasets show that the proposed method outperforms the state-of-the-art baselines in the cross-modal retrieval tasks.
Keywords:
Cross-modal retrieval
deep supervised hashing
margin-dynamic-softmax loss
proxy code

Journal

IEEE Transactions on Knowledge and Data Engineering cover
IEEE Transactions on Knowledge and Data Engineering
IF:
10.4
Papers:
6.8K
Citations:
3.2W

Organization

S
Sun Yat Sen University
Scholars:
9.9W
Papers: 7.2W
Citations: 95
B
beijing institute of technology
Scholars:
5.5W
Papers: 4.0W
Citations: 63
Z
zhejiang university
Scholars:
17.7W
Papers: 12.1W
Citations: 152
researcher View more organizations
Cited Papers

Cited Papers

errShare
errSave
Is There a Role for Magnetic Resonance Imaging in Diagnosing Colovesical Fistulas?
err2008-10-01
err0
PREAI
errS. Ravichandran; H.U. Ahmed; S.S. Matanhelia; M. Dobson
errShare
errSave
Deep Cross-Modal Hashing With Hashing Functions and Unified Hash Codes Jointly Learning
err2022-02-01
err45
errOAAI
errTu, Rong-Cheng; Mao, Xian-Ling; Ma, Bing; Hu, Yong; Yan, Tan; Wei, Wei; Huang, Heyan
errShare
errSave
Learning Discriminative Binary Codes for Large-scale Cross-modal Retrieval
err2017-05-01
err382
PREAI
errXu, Xing; Shen, Fumin; Yang, Yang; Shen, Heng Tao; Li, Xuelong
errShare
errSave
Multi-Level Policy and Reward-Based Deep Reinforcement Learning Framework for Image Captioning
err2020-05-01
err78
PREAI
errXu, Ning; Zhang, Hanwang; Liu, An-An; Nie, Weizhi; Su, Yuting; Nie, Jie; Zhang, Yongdong
errShare
errSave
MANAGEMENT OF RENAL TRAUMA A RET R OSPECTIVE STUDY - OUR EXPERIENCE
err2015-10-21
err0
errOAAI
errPrakash Babu S M L; Sandeep Puvvada; Avinash Patil; Arvind Nayak; Nagaraj H K
errShare
errSave
researcher View more