arrow
返回

EdVAE: Mitigating codebook collapse with evidential discrete variational autoencoders

delete2024-12-01
delete0
delete
OA
AI
G
Gulcin Baykal *
M
Melih Kandemir
G
Gözde Ünal
DOI:10.1016/j.patcog.2024.110792delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Codebook collapse is a common problem in training deep generative models with discrete representation spaces like Vector Quantized Variational Autoencoders (VQ-VAEs). We observe that the same problem arises for the alternatively designed discrete variational autoencoders (dVAEs) whose encoder directly learns a distribution over the codebook embeddings to represent the data. We hypothesize that using the softmax function to obtain a probability distribution causes the codebook collapse by assigning overconfident probabilities to the best matching codebook elements. In this paper, we propose a novel way to incorporate evidential deep learning (EDL) through a hierarchical Bayesian modeling instead of softmax to combat the codebook collapse problem of dVAE. We evidentially monitor the significance of attaining the probability distribution over the codebook embeddings, in contrast to softmax usage. Our experiments using various datasets show that our model, called EdVAE, mitigates codebook collapse while improving the reconstruction performance, and enhances the codebook usage compared to dVAE and VQ-VAE based models. Our code can be found at https://github.com/ituvisionlab/EdVAE.
Keyword:
Vector quantized variational autoencoders
Discrete variational autoencoders
Evidential deep learning
Codebook collapse
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Pattern Recognition 封面图
Pattern Recognition
IF:
7.6
论文数:
1.3W
被引数:
4.5W

机构

U
University of Southern Denmark
学者数:
2.1W
论文数: 2.0W
被引数: 2.9W
I
Istanbul Technical University
学者数:
8.9K
论文数: 7.8K
被引数: 7.9K
引用论文

引用论文

Uncertainty estimation for stereo matching based on evidential deep learning基于证据深度学习的立体匹配不确定性估计
err2022-04-01
err120
errOAAI
errWang, Chen; Wang, Xiang; Zhang, Jiawei; Zhang, Liang; Bai, Xiao; Ning, Xin; Zhou, Jun; Hancock, Edwin
err分享
err收藏
Reparameterizing and dynamically quantizing image features for image generation
err2024-02-01
err4
PREAI
errSun, Mingzhen; Wang, Weining; Zhu, Xinxin; Liu, Jing
err分享
err收藏