arrow
返回

Learning MLatent Representations for Generalized Zero-Shot Learning

delete2023-01-01
delete6
PRE
AI
Y
Yalan Ye
T
Tongjie Pan
T
Tonghoujun Luo
J
Jingjing Li *
H
Heng Tao Shen
DOI:10.1109/TMM.2022.3145237delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
In generative adversarial network (GAN) based zero-shot learning (ZSL) approaches, the synthesized unseen visual features are inevitably prone to seen classes since the feature generator is merely trained on seen references, which causes the inconsistency between visual features and their corresponding semantic attributes. This visual-semantic inconsistency is primarily induced by the non-preserved semantic-relevant components and the non-rectified semantic-irrelevant low-level visual details. Existing generative models generally tackle the issue by aligning the distribution of the two modalities with an additional visual-to-semantic embedding, which tends to cause the hubness problem and ruin the diversity of visual modality. In this paper, we propose a novel generative model named learning modality-consistent latent representations GAN (LCR-GAN) to address the problem via embedding the visual features and their semantic attributes into a shared latent space. Specifically, to preserve the semantic-relevant components, the distributions of the two modalities are aligned by maximizing the mutual information between them. And to rectify the semantic-irrelevant visual details, the mutual information between original visual features and their latent representations is confined within an appropriate range. Meanwhile, the latent representations are decoded back to both modalities to further preserve the semantic-relevant components. Extensive evaluations on four public ZSL benchmarks validate the superiority of our method over other state-of-the-art methods.
Keyword:
Visualization
Semantics
Mutual information
Image color analysis
Generative adversarial networks
Task analysis
Cows
Generative adversarial network
latent representations
mutual information
semantic-relevant
semantic-irrelevant
zero-shot learning

期刊

IEEE Transactions on Multimedia 封面图
IEEE Transactions on Multimedia
IF:
9.7
论文数:
4.5K
被引数:
2.4W

机构

暂无机构信息
引用论文

引用论文

Sclerostin deficient mice rapidly heal bone defects by activating β-catenin and increasing intramembranous ossification
err2013-11-01
err0
errOAAI
errMeghan E. McGee-Lawrence; Zachary C. Ryan; Lomeli R. Carpio; Sanjeev Kakar; Jennifer J. Westendorf; Rajiv Kumar
err分享
err收藏
Design and realization of a mobile wheelchair robot for all terrains
err2012-04-02
err0
PREAI
errChun-Ta Chen; Chieh-Chuan Feng; Yu-An Hsieh
err分享
err收藏
err分享
err收藏
Direct fabrication of rigid microstructures on a metallic roller using a dry film resist
err2007-11-28
err0
PREAI
errLiang-Ting Jiang; Tzu-Chien Huang; Chih-Yuan Chang; Jian-Ren Ciou; Sen-Yeu Yang; Po-Hsun Huang
err分享
err收藏
err分享
err收藏
学者 查看更多内容