arrow
返回

Diffpvt:information filtering based diffusion model with PVT for medical image segmentation

delete2025-01-04
delete0
PRE
AI
C
Chengming Wang
G
Genji Yuan
M
Mengjun Li *
李晋江 封面图
李晋江 (Jinjiang Li)
DOI:10.1007/s13042-024-02519-3delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Diffusion probabilistic model, as a generative model, has gained wide attention in image processing tasks, and its powerful image generation capability has made it flourish in the field of computer vision. In recent years, it has achieved excellent performance in the field of medical image segmentation, and nowadays denoising diffusion models supported by network architectures such as U-Net and Transformer are all successfully applied in medical image segmentation tasks. In order to explore the potential of pyramid structure in diffusion modelling, we propose a denoising diffusion model, DiffPVT, which uses the PVT as the network architecture. In order to ensure the diffusion model's compatibility with the pyramid vision transformer (PVT), we use a dual-encoder architecture to convey the spatial information to ensure the effectiveness of the diffusion model in the denoising stage, while proposing a novel decoder structure to further enrich the reconstruction process, in which we retain the information embedding of the time steps in the second layer encoding stage while applying its extension to the decoder part. We combine the denoising task of the diffusion model with the critical need of the edge detection task in the field of medical image segmentation, and propose dual-mode information filtering module (DMIFM) and use it as a holistic structure to process the requirements of multiple types of tasks in parallel, thus enhancing the denoising process of the diffusion model and enriching the edge information of the feature images. We conduct extensive experiments on four public datasets and confirm that DiffPVT has excellent segmentation level in the field of medical image segmentation. https://github.com/cn-xvkong/DiffPVT.
Keyword:
Denoising diffusion model
Pyramid vision transformer
Dual-mode information filtering
Squeeze-excitation residual fusion decoder
Medical image segmentation

期刊

International Journal of Machine Learning and Cybernetics 封面图
International Journal of Machine Learning and Cybernetics
IF:
2.7
论文数:
3.2K
被引数:
5.6K

机构

暂无机构信息
引用论文

引用论文

Risk Factors for Running-Related Injuries in Trailrunners
err2016-05-01
err0
PREAI
errLuiz Carlos Hespanhol; Willem van Mechelen; Evert Verhagen
err分享
err收藏
SILP: Enhancing skin lesion classification with spatial interaction and local perception
err2024-12-01
err0
PREAI
errNguyen, Khanh-Duy; Zhou, Yu-Hui; Nguyen, Quoc-Viet; Sun, Min-Te; Sakai, Kazuya; Ku, Wei-Shinn
err分享
err收藏
Gland segmentation in colon histology images: The glas challenge contest结肠组织学图像中的腺体分割: gla挑战赛
err2017-01-01
err495
errOAAI
errSirinukunwattana, Korsuk; Pluim, Josien P. W.; Chen, Hao; Qi, Xiaojuan; Heng, Pheng-Ann; Guo, Yun Bo; Wang, Li Yang; Matuszewski, Bogdan J.; Bruni, Elia; Sanchez, Urko; Bohm, Anton; Ronneberger, Olaf; Cheikh, Bassem Ben; Racoceanu, Daniel; Kainz, Philipp; Pfeiffer, Michael; Urschler, Martin; Snead, David R. J.; Rajpoot, Nasir M.
err分享
err收藏
err分享
err收藏
Introduction
err2002-01-01
err0
PREAI
errJoanne Neale
err分享
err收藏
Validation of Questionnaire and Diary Measures of Time Outdoors Against an Objective Measure of Personal Ultraviolet Radiation Exposure
err2018-03-25
err0
errOAAI
errAnne E Cust; Georgina L Fenton; Amelia K Smit; David Espinoza; Suzanne Dobbinson; Alison Brodie; Huong Tran Cam Dang; Michael G Kimlin
err分享
err收藏
Metal Organic Framework-Based Dispersive Solid-Phase Microextraction of Carbaryl from Food and Water Prior to Detection by Ultra-Performance Liquid Chromatography-Tandem Mass Spectrometry
err2022-01-28
err0
errOAAI
errMohamed A. Habila; Bushra Alhenaki; Adel El-Marghany; Mohamed Sheikh; Ayman A. Ghfar; Zeid A. ALOthman; Mustafa Soylak
err分享
err收藏
学者 查看更多内容