arrow
返回

Semantic-Interactive Graph Convolutional Network for Multilabel Image Recognition

delete2022-08-01
delete17
PRE
AI
B
Bingzhi Chen
张政 封面图
张政 (Zheng Zhang)
卢瑶 (Yao Lu) *
F
Fanglin Chen
G
Guangming Lu
章典 封面图
章典 (David Zhang)
DOI:10.1109/TSMC.2021.3103842delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Multilabel image recognition, a critically practical task in computer vision, aims to predict multiple objects present in each image. The existing studies mainly focus on conceptual visual cues but fail to reconcile the visual information with their semantic guidance. Intuitively, humans can not only associate extra topological concepts but also imagine other approximate scenes based on a semantic description. Inspired by such semantic-interactive capability, two different types of semantic priors, i.e., the concept correlations of the same scene and semantic similarities among different scenes, should be further explored for the recognition decisions. To efficiently interact with these semantic relationships, in this article, we propose a novel semantic-interactive graph convolutional network (SI-GCN), which can leverage the topological information learned from knowledge graphs to boost the performance of multilabel recognition. Specifically, the proposed SI-GCN framework consists of two different GCN-based branches in parallel, i.e., concept correlations learning (CCL) branch and semantic similarity learning (SSL) branch. Inputting the semantic-embedding vectors of all the concepts, the CCL branch maps the label co-occurrence graph into a set of interdependent concept classifiers. Recalibrating the image feature embedding with the standardized supervision of the semantic similarity graph, the SSL branch learns the semantically consistent in-batch visual representations. Finally, a well-established interactive learning scheme is formulated to concurrently optimize the obtained concept classifiers and the visual representation learning in an end-to-end manner. Extensive experiments on the MS-COCO and Pascal VOC 2007 & 2012 benchmarks demonstrate the superiorities of the proposed SI-GCN method compared to the state-of-the-art baselines.
Keyword:
Semantics
Image recognition
Visualization
Correlation
Feature extraction
Task analysis
Training
Graph convolutional network (GCN)
label co-occurrence
multilabel image recognition
semantic interactive
semantic similarity

期刊

IEEE Transactions on Cybernetics 封面图
IEEE Transactions on Cybernetics
IF:
10.5
论文数:
1.1W
被引数:
5.0W

机构

H
harbin institute of technology
学者数:
8.0W
论文数: 6.6W
被引数: 66
T
The Chinese University of Hong Kong, Shenzhen
学者数:
4.3K
论文数: 4.0K
被引数: 7
引用论文

引用论文

err分享
err收藏
err分享
err收藏
Absorption spectroscopy of surfactant-dispersed carbon nanotube film: Modulation of electronic structures
err2008-04-01
err0
PREAI
errHong-Zhang Geng; Dae Sik Lee; Ki Kang Kim; Gang Hee Han; Hyeon Ki Park; Young Hee Lee
err分享
err收藏
A Multi-Degree of Freedom Tuned Mass Damper Design for Vibration Mitigation of a Suspension Bridge
err2020-01-08
err0
errOAAI
errFanhao Meng; Jiancheng Wan; Yongjun Xia; Yong Ma; Jingjun Yu
err分享
err收藏
err分享
err收藏
学者 查看更多内容