返回
MDFN: Multi-scale deep feature learning network for object detection
DOI:10.1016/j.patcog.2019.107149.png)
摘要
En 中文
This paper proposes an innovative object detector by leveraging deep features learned in high-level layers. Compared with features produced in earlier layers, the deep features are better at expressing semantic and contextual information. The proposed deep feature learning scheme shifts the focus from concrete features with details to abstract ones with semantic information. It considers not only individual objects and local contexts but also their relationships by building a multi-scale deep feature learning network (MDFN). MDFN efficiently detects the objects by introducing information square and cubic inception modules into the high-level layers, which employs parameter-sharing to enhance the computational efficiency. MDFN provides a multi-scale object detector by integrating multi-box, multi-scale and multi-level technologies. Although MDFN employs a simple framework with a relatively small base network (VGG-16), it achieves better or competitive detection results than those with a macro hierarchical structure that is either very deep or very wide for stronger ability of feature extraction. The proposed technique is evaluated extensively on KITTI, PASCAL VOC, and COCO datasets, which achieves the best results on KITTI and leading performance on PASCAL VOC and COCO. This study reveals that deep features provide prominent semantic information and a variety of contextual contents, which contribute to its superior performance in detecting small or occluded objects. In addition, the MDFN model is computationally efficient, making a good trade-off between the accuracy and speed. (C) 2019 Elsevier Ltd. All rights reserved.
Keyword:
Deep feature learning
Multi-scale
Semantic and contextual information
Small and occluded objects
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
7.6
论文数:
1.3W
被引数:
4.5W
机构
引用论文
Scene and place recognition using a hierarchical latent topic model使用分层潜在主题模型进行场景和地点识别
NEUROCOMPUTING
IF6.5
Dictionary Representation of Deep Features for Occlusion-Robust Face Recognition用于遮挡鲁棒人脸识别的深度特征字典表示
IEEE ACCESS
IF3.6
Gradient-based learning applied to document recognition基于梯度的学习在文档识别中的应用
PROCEEDINGS OF THE IEEE
IF25.9
Toward learning a unified many-to-many mapping for diverse image translation
PATTERN RECOGNITION
IF7.6

