arrow
Return

Video diffusion generation: comprehensive review and open problems

delete2025-08-20
delete0
delete
OA
AI
W
Wenping Ma
X
Xiaoting Yang
L
Licheng Jiao *
L
Lingling Li
X
Xu Liu
刘芳 cover
刘芳 (Fang Liu)
陈璞花 (Puhua Chen)
Y
Yuting Yang
M
Mengru Ma
L
Long Sun
R
Ruohan Zhang
X
Xueli Geng
Y
Yuwei Guo
S
Shuyuan Yang
Z
Zhixi Feng
DOI:10.1007/s10462-025-11331-6delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Video generation has become an increasingly important component of AI-generated content (AIGC), owing to its rich semantic expressiveness and growing application potential. Among various generative paradigms, diffusion models have recently gained prominence due to their strong controllability, competitive visual quality, and compatibility with multimodal inputs. However, most existing surveys provide limited coverage of diffusion-based video generation, often lacking systematic analysis and comprehensive comparisons. To address this gap, this paper presents a thorough and structured review of diffusion models for video generation. We first outline the theoretical foundations and core architectures of diffusion models, and then the key design principles of representative methods for video generation were introduced. We propose a unified taxonomy that categorizes over two hundred methods, analyzing their key characteristics, strengths, and limitations. In addition, we compared the performance of classical methods and summarized commonly used datasets and evaluation metrics in this field for ease of model benchmarking and selection. Finally, we discuss open problems and future research directions, aiming to provide a valuable reference for both academic research and practical development.
Keywords:
Review
AI-generated content (AIGC)
Video diffusion models
Video generation

Journal

Artificial Intelligence Review cover
Artificial Intelligence Review
IF:
13.9
Papers:
6.1K
Citations:
1.9W

Organization