arrow
返回

Unsupervised Object Transfiguration with Attention

delete2019-04-08
delete6
PRE
AI
Z
Zihan Ye
F
Fan Lyu
L
Linyan Li *
Y
Yu Sun
Q
Qiming Fu
胡
胡伏原 (Fuyuan Hu) *
DOI:10.1007/s12559-019-09633-3delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Object transfiguration is a subtask of the image-to-image translation, which translates two independent image sets and has a wide range of applications. Recently, some studies based on Generative Adversarial Network (GAN) have achieved impressive results in the image-to-image translation. However, the object transfiguration task only translates regions containing target objects instead of whole images; most of the existing methods never consider this issue, which results in mistranslation on the backgrounds of images. To address this problem, we present a novel pipeline called Deep Attention Unit Generative Adversarial Networks (DAU-GAN). During the translating process, the DAU computes attention masks that point out where the target objects are. DAU makes GAN concentrate on translating target objects while ignoring meaningless backgrounds. Additionally, we construct an attention-consistent loss and a background-consistent loss to compel our model to translate intently target objects and preserve backgrounds further effectively. We have comparison experiments on three popular related datasets, demonstrating that the DAU-GAN achieves superior performance to the state-of-the-art. We also export attention masks in different stages to confirm its effect during the object transfiguration task. The proposed DAU-GAN can translate object effectively as well as preserve backgrounds information at the same time. In our model, DAU learns to focus on the most important information by producing attention masks. These masks compel DAU-GAN to effectively distinguish target objects and backgrounds during the translation process and to achieve impressive translation results in two subsets of ImageNet and CelebA. Moreover, the results show that we cannot only investigate the model from the image itself but also research from other modal information.
Keyword:
Multi-modalities
Object transfiguration
Image-to-image translation
Generative Adversarial Networks (GANs)
Attention mechanism
Deep learning
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Cognitive Computation 封面图
Cognitive Computation
IF:
4.3
论文数:
1.6K
被引数:
3.6K

机构

T
tianjin university
学者数:
8.0W
论文数: 5.8W
被引数: 88
S
suzhou institute of trade & commerce (sitc)
学者数:
29
论文数: 21
被引数: 0
S
suzhou university of science & technology
学者数:
5.0K
论文数: 4.8K
被引数: 4
学者 查看更多机构
引用论文

引用论文

Visual Attention Model Based Vehicle Target Detection in Synthetic Aperture Radar Images: A Novel Approach
err2014-12-06
err24
PREAI
errGao, Fei; Zhang, Ye; Wang, Jun; Sun, Jinping; Yang, Erfu; Hussain, Amir
err分享
err收藏
Zero-Shot Learning via Attribute Regression and Class Prototype Rectification
err2018-02-01
err51
PREAI
errLuo, Changzhi; Li, Zhetao; Huang, Kaizhu; Feng, Jiashi; Wang, Meng
err分享
err收藏
Unsupervised image saliency detection with Gestalt-laws guided optimization and visual attention based refinement
err2018-07-01
err138
errOAAI
errYan, Yijun; Ren, Jinchang; Sun, Genyun; Zhao, Huimin; Han, Junwei; Li, Xuelong; Marshall, Stephen; Zhan, Jin
err分享
err收藏
Visual Attribute Transfer through Deep Image Analogy
err2017-07-20
err245
errOAAI
errLiao, Jing; Yao, Yuan; Yuan, Lu; Hua, Gang; Kang, Sing Bing
err分享
err收藏
Cognitive Fusion of Thermal and Visible Imagery for Effective Detection and Tracking of Pedestrians in Videos
err2017-12-04
err60
errOAAI
errYan, Yijun; Ren, Jinchang; Zhao, Huimin; Sun, Genyun; Wang, Zheng; Zheng, Jiangbin; Marshall, Stephen; Soraghan, John
err分享
err收藏
学者 查看更多内容