arrow
返回

Continuous frame motion sensitive self-supervised collaborative network for video representation learning

delete2023-04-01
delete5
PRE
AI
胡
胡正平 (Zhengping Hu) *
赵梦瑶 封面图
赵梦瑶 (Mengyao Zhao)
H
Hehao Zhang
Z
Zhe Sun
DOI:10.1016/j.aei.2023.101941delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Motion, as a feature of video that changes in temporal sequences, is crucial to visual understanding. The powerful video representation and extraction models are typically able to focus attention on motion features in challenging dynamic environments to complete more complex video understanding tasks. However, previous approaches discriminate mainly based on similar features in the spatial or temporal domain, ignoring the interdependence of consecutive video frames. In this paper, we propose the motion sensitive self-supervised collaborative network, a video representation learning framework that exploits a pretext task to assist feature comparison and strengthen the spatiotemporal discrimination power of the model. Specifically, we first propose the motion-aware module, which extracts consecutive motion features from the spatial regions by frame difference. The global-local contrastive module is then introduced, with context and enhanced video snippets being defined as appropriate positive samples for a broader feature similarity comparison. Finally, we introduce the snippet operation prediction module, which further assists contrastive learning to obtain more reliable global semantics by sensing changes in continuous frame features. Experimental results demonstrate that our work can effectively extract robust motion features and achieve competitive performance compared with other state-of-the-art self-supervised methods on downstream action recognition and video retrieval tasks.
Keyword:
Self-supervised representation learning
Pretext task
Global-local contrastive learning
Action recognition
Video retrieval

期刊

Advanced Engineering Informatics 封面图
Advanced Engineering Informatics
IF:
9.9
论文数:
4.4K
被引数:
1.7W

机构

Y
Yanshan University
学者数:
1.7W
论文数: 1.1W
被引数: 1.3W
引用论文

引用论文

err
IF0
err
err0
PREAI
err
err分享
err收藏
err分享
err收藏
Diagnostic ability of confocal near-infrared reflectance fundus imaging to detect retrograde microcystic maculopathy from chiasm compression. A comparative study with OCT findings
err2021-06-24
err0
errOAAI
errMário L. R. Monteiro; Rafael M. Sousa; Rafael B. Araújo; Daniel Ferraz; Mohammad A. Sadiq; Leandro C. Zacharias; Rony C. Preti; Leonardo P. Cunha; Quan D. Nguyen
err分享
err收藏
Climate policy: Steps to China's carbon peak气候政策: 迈向中国碳峰值的步骤
err2015-06-17
err0
errOAAI
errZhu Liu; Dabo Guan; Scott Moore; Henry Lee; Jun Su; Qiang Zhang
err分享
err收藏
Strain-tunable electronic, elastic, and optical properties of CaI2monolayer: first-principles study
err2020-04-17
err0
PREAI
errXiao-Fang Chen; Li Wang; Zhao-Yi Zeng; Xiang-Rong Chen; Qi-Feng Chen
err分享
err收藏
err分享
err收藏
学者 查看更多内容