arrow
Return

Multi-scale global-local collaborative learning for accurate significant object detection

delete2026-08-11
delete0
delete
OA
AI
J
Jiayin Liu
W
Wang Zetong
S
Shen Yuyuan
S
Shen Teng
L
Lu Liang
Z
Zhang Yongqiang
L
Liu Zhi
H
Huilin Jiang *
DOI:10.1038/s41598-026-63649-ydelete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Salient Object Detection (SOD) remains a fundamental task in computer vision and visual computing, supporting applications ranging from image understanding to human-computer interaction. Existing methods still face two coupled challenges: insufficient modeling of multi-scale salient structures and imbalanced fusion between global semantic information and local details, which often lead to incomplete salient regions and blurred boundaries. To address these issues, this study proposes MSGAN, a multi-scale global-local collaborative learning framework that integrates multi-scale mixed convolution and adaptive global-local attention to enhance feature representation. Extensive experiments on the HKU-IS, ECSSD, PASCAL-S, and DUT-OMRON datasets demonstrate that our method achieves significant improvements in F-measure, MAE, and Em metrics, outperforming state-of-the-art approaches. Ablation studies validate the effectiveness of each core component. This work advances robust SOD for complex real-world scenarios and provides insights into attention-guided visual perception.
Keywords:
Salient object detection
Multi-scale feature extraction
Global-local attention
Encoder-decoder architecture
Feature fusion
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Scientific Reports cover
Scientific Reports
IF:
3.9
Papers:
27.1W
Citations:
83.5W

Organization

A
Aviation Industry Corporation of China
Scholars:
62
Papers: 38
Citations: 400
C
College of Optoelectronic Engineering
Scholars:
25
Papers: 10
Citations: 0