arrow
Return

Enhancing object detection with large kernel convolution and cross convolution

delete2025-06-25
delete0
PRE
AI
Y
Yaqian Li
G
Guoping Liu
H
Haibin Li
W
Wenming Zhang
X
Xiaoyang Shen
DOI:10.1016/j.dsp.2025.105433delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Existing object detection models often struggle with detecting small objects due to their limited ability to capture sufficient contextual information. In this paper, we introduce a lightweight object detection model that leverages large kernel convolution with attention (LKA) and a hierarchical feature fusion group (HFFG) to address this issue. The LKA module employs large kernel convolution to capture long-range dependencies and contextual information, combined with depthwise separate convolution to maintain a lightweight design. An incorporated attention mechanism further enables the modal to adaptively focus on key areas, thereby improving detection performance for small objects. The HFFG module, which integrates Cross Convolution Blocks, explores and retains structural information across different scales. By effectively extracting structural details, our model exhibits enhanced performance on object of various sizes. Extensive experiments on the VisDrone2019 and PASACAL VOC datasets demonstrate that our model achieves an outstanding mAP of 23.4 %, surpassing the baseline YOLOX-s model by +1.5 %. These results not only validate the effectiveness but also demonstrate its robustness and generalization capability.

Journal

Signal Processing cover
Signal Processing
IF:
3.6
Papers:
10.0K
Citations:
1.7W

Organization

No organization information available
Cited Papers

Cited Papers

Dehazing & Reasoning YOLO: Prior knowledge-guided network for object detection in foggy weather
err2024-12-01
err1
PREAI
errZhong, Fujin; Shen, Wenxin; Yu, Hong; Wang, Guoyin; Hu, Jun
errShare
errSave
Edge guidance filtering for structure extraction
err2022-09-02
err3
PREAI
errSun, Beichen; Qi, Yuehan; Zhang, Guanyu; Liu, Yang
errShare
errSave
The Pascal Visual Object Classes (VOC) Challenge
err2009-09-09
err9.0K
PREAI
errEveringham, Mark; Van Gool, Luc; Williams, Christopher K. I.; Winn, John; Zisserman, Andrew
errShare
errSave
errShare
errSave
Going deeper with convolutions
err2015-06-01
err0
errOAAI
errChristian Szegedy; Wei Liu; Yangqing Jia; Pierre Sermanet; Scott Reed; Dragomir Anguelov; Dumitru Erhan; Vincent Vanhoucke; Andrew Rabinovich
errShare
errSave
Visual attention network
err2023-12-01
err224
errOAAI
errGuo, Meng-Hao; Lu, Cheng-Ze; Liu, Zheng-Ning; Cheng, Ming-Ming; Hu, Shi-Min
errShare
errSave
no more