arrow
返回

Cross-modal semantic autoencoder with embedding consensus

delete2021-10-13
delete1
delete
OA
AI
S
Shengzi Sun
Z
Zhilong Mi
Z
Zhiming Zheng
DOI:10.1038/s41598-021-92750-7delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Cross-modal retrieval has become a topic of popularity, since multi-data is heterogeneous and the similarities between different forms of information are worthy of attention. Traditional single-modal methods reconstruct the original information and lack of considering the semantic similarity between different data. In this work, a cross-modal semantic autoencoder with embedding consensus (CSAEC) is proposed, mapping the original data to a low-dimensional shared space to retain semantic information. Considering the similarity between the modalities, an automatic encoder is utilized to associate the feature projection to the semantic code vector. In addition, regularization and sparse constraints are applied to low-dimensional matrices to balance reconstruction errors. The high dimensional data is transformed into semantic code vector. Different models are constrained by parameters to achieve denoising. The experiments on four multi-modal data sets show that the query results are improved and effective cross-modal retrieval is achieved. Further, CSAEC can also be applied to fields related to computer and network such as deep and subspace learning. The model breaks through the obstacles in traditional methods, using deep learning methods innovatively to convert multi-modal data into abstract expression, which can get better accuracy and achieve better results in recognition.
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Scientific Reports 封面图
Scientific Reports
IF:
3.9
论文数:
28.0W
被引数:
83.5W

机构

B
Beihang University
学者数:
5.2W
论文数: 4.1W
被引数: 37
引用论文

引用论文

Tension control: dancer rolls or load cells
err1993-01-01
err0
PREAI
errN.A. Ebler; R. Arnason; G. Michaelis; N. D'Sa
err分享
err收藏
err分享
err收藏
err分享
err收藏
学者 查看更多内容