arrow
返回

Y-Autoencoders: Disentangling latent representations via sequential encoding

delete2020-12-01
delete10
delete
OA
AI
M
Massimiliano Patacchiola *
P
Patrick Fox‐Roberts
E
Edward Rosten
DOI:10.1016/j.patrec.2020.09.025delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
In the last few years there have been important advancements in disentangling latent representations using generative models, with the two dominant approaches being Generative Adversarial Networks (GANs) and Variational Autoencoders (VAEs). However, standard Autoencoders (AEs) and closely related structures have remained popular because they are easy to train and adapt to different tasks. An interesting question is if we can achieve state-of-the-art latent disentanglement with AEs while retaining their good properties. We propose an answer to this question by introducing a new model called Y-Autoencoder (Y-AE). The structure and training procedure of a Y-AE enclose a representation into an implicit and an explicit part. The implicit part is similar to the output of an AE and the explicit part is strongly correlated with labels in the training set. The two parts are separated in the latent space by splitting the output of the encoder into two paths (forming a Y shape) before decoding and re-encoding. We then impose a number of losses, such as reconstruction loss, and a loss on dependence between the implicit and explicit parts. Additionally, the projection in the explicit manifold is monitored by a predictor, that is embedded in the encoder and trained end-to-end with no adversarial losses. We provide significant experimental results on various domains, such as separation of style and content, image-to-image translation, and inverse graphics. (C) 2020 Elsevier B.V. All rights reserved.
Keyword:
Disentangled representations
Deep learning
Autoencoders
Generative models
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Pattern Recognition Letters 封面图
Pattern Recognition Letters
IF:
3.3
论文数:
8.0K
被引数:
1.6W

机构

U
University of Edinburgh
学者数:
5.2W
论文数: 4.6W
被引数: 71
引用论文

引用论文

Multimodal Deep Autoencoder for Human Pose Recovery
err2015-12-01
err520
PREAI
errHong, Chaoqun; Yu, Jun; Wan, Jian; Tao, Dacheng; Wang, Meng
err分享
err收藏
err分享
err收藏
Deep generative video prediction
err2018-07-01
err13
PREAI
errYu, Tingzhao; Wang, Lingfeng; Gu, Huxiang; Xiang, Shiming; Pan, Chunhong
err分享
err收藏
err分享
err收藏
err分享
err收藏
没有更多内容