arrow
Return

Deblurring with Improved Video Diffusion Model

delete2026-01-01
delete0
PRE
AI
H
Haoyang Long *
Z
Zheng, Bo
W
Wufan Wang
张政 cover
张政 (Zheng Zhang)
王文东 cover
王文东 (Wendong Wang)
DOI:10.1007/978-3-032-04546-1_5delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Video deblurring faces significant challenges due to complex blur patterns arising from both camera motion and object movement. While existing methods predominantly rely on distortion-based metrics like PSNR for evaluation, these measurements often show limited correlation with human visual perception and tend to produce unrealistic reconstructions. Recent advancements in diffusion models have demonstrated exceptional capabilities in generating realistic visual content, with image diffusion models achieving photorealistic synthesis and video diffusion models showing promise in temporal coherence. In this paper, we propose DIVD: Deblurring with Improved Video Diffusion Model, specifically designed for deblurring tasks. Our framework introduces two key enhancements: 1) Window-based Temporal Self-Attention (WTSA) for leveraging inter-frame dependencies. 2) Multi-frame Relative Positional Encoding (MRPE) for handling inter-frame inconsistencies. Extensive experiments demonstrate our model's state-of-the-art performance across multiple perceptual metrics while maintaining competitive distortion metrics. Qualitative evaluations reveal significantly improved detail preservation compared to existing approaches. This work presents successful adaptation of diffusion models for video deblurring, effectively addressing the realism limitations of conventional approaches.
Keywords:
Video Deblurring
Diffusion Model
Temporal Coherence

Journal

A
ARTIFICIAL NEURAL NETWORKS AND MACHINE LEARNING-ICANN 2025, PT II
IF:
0
Papers:
42
Citations:
0

Organization

B
beijing university of posts & telecommunications
Scholars:
1.4W
Papers: 1.2W
Citations: 9