arrow
返回

A knowledge-guided multimodal network for video summarization

delete2026-04-23
delete0
PRE
AI
X
Xiaoyan Tian
Y
Ye Jin *
Z
Zhao Zhang
刘鹏 封面图
刘鹏 (Peng Liu)
F
Fenglei Ni
DOI:10.1016/j.eswa.2026.132566delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
• 基于知识的编码器提供高级隐式知识以增强视频摘要。 • 设计了一个从精细到粗略的空间投影模块来建模多粒度多模态表示。 • 预测头保持同一镜头内语义的完整性和平滑性。 • 实验结果证明了所提出模型的有效性。
Keyword:
knowledge-guided
multimodal network
video summarization
representation learning
shot integrity

期刊

Expert Systems with Applications 封面图
Expert Systems with Applications
IF:
7.5
论文数:
3.0W
被引数:
10.2W

机构

H
Harbin Institute of Technology
学者数:
1.6W
论文数: 5.0K
被引数: 8.5W
引用论文

引用论文

暂无论文信息