arrow
Return

A cross frame post-processing strategy for video object detection

delete2022-07-01
delete4
PRE
AI
X
Xin Song *
Z
Ziqiang Qi
朱剑林 (Jianlin Zhu)
S
Shuhua Li
DOI:10.1016/j.displa.2022.102230delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Video-based object detection plays an important role in the real world and scientific research. Compared with still images, video detection is more challenging due to occlusion, rare poses, high-speed movement, frames loss, etc. In order to improve the existing video stream detectors widely and with low coupling, a post-processing strategy, CFPP, is proposed in this work. The framework can establish a cross frame link based on deep learning, connect the proposals belonging to the same object, and improve the performance of the detector by optimizing the classification confidence and object coordinates. Furthermore, CFPP can connect the proposals in adjacent and non adjacent frames at the same time, which makes it exploit the context information of video stream more effectively than other post-processing strategies. Experiments shows that CFPP can improve the existing detectors (e.g. we improve the mAP of YOLOv4 on ImageNet VID dataset form 69.24% to 78.15%). In addition, experiments show that the designed framework can achieve better detection effect than other strategies in the case of high-speed moving object and frames loss.
Keywords:
Video object detection
Post-processing
Deep learning
Optimization algorithm

Journal

Displays cover
Displays
IF:
3.4
Papers:
2.1K
Citations:
3.2K

Organization

N
northeastern university - china
Scholars:
3.1W
Papers: 2.7W
Citations: 37