arrow
Return

Human-Object interaction detection algorithm based on visual-semantic interaction perception

delete2025-09-22
delete0
PRE
AI
Q
Qing Ye *
D
Deng, Gaohui
X
Xiuju Xu
Y
Yongmei Zhang
DOI:10.1007/s11760-025-04804-2delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
To address the challenges of inadequate extraction of features from interaction regions due to occlusion and the limited generalization in small-sample training, we propose a human-object interaction detection algorithm based on visual-semantic interaction perception. To improve the extraction of interaction region features hindered by occlusion, we propose a Deep Visual Interaction Feature Extraction Network(DVIF). Using graph convolution to develop an adaptable human-object relationship graph, paired with an optimized human pose segment, allows us to better capture significant interaction region characteristics. To enhance the generalization ability of small sample training data, a Semantic-driven Sample Expansion Module (SSE) is proposed. This module incorporates word embeddings and linguistic priors into data generation to augment the training samples. We propose a Cross-modal External Attention Fusion Module (CEA) to synergistically enhance visual and semantic information, thereby improving the accuracy and robustness of interaction detection. Experiments on the HICO-DET and V-COCO datasets achieved human-object interaction detection accuracy of 30.67% and 60.3%, demonstrating the effectiveness of the proposed algorithm.
Keywords:
Human-Object interaction detection algorithm
Deep visual interaction feature extraction
Graph convolutional network
Pose estimation
Semantic-driven sample expansion
Cross-modal external attention fusion

Journal

Signal Image and Video Processing cover
Signal Image and Video Processing
IF:
2.1
Papers:
908
Citations:
4.6K

Organization

No organization information available
Cited Papers

Cited Papers

Mask R-CNN
err2017-10-01
err0
PREAI
errKaiming He; Georgia Gkioxari; Piotr Dollar; Ross Girshick
errShare
errSave
Learning to Detect Human-Object Interactions
err2018-03-01
err0
errOAAI
errYu-Wei Chao; Yunfan Liu; Xieyang Liu; Huayi Zeng; Jia Deng
errShare
errSave
researcher View more