返回
Joint Input and Output Space Learning for Multi-Label Image Classification
DOI:10.1109/TMM.2020.3002185.png)
摘要
En 中文
Multi-label image classification aims to predict the labels associated with a given image. While most existing methods utilize unified image representations, extracting label-specific features through input space learning would improve the discriminative power of the learned features. On the other hand, most feature learning studies often ignore the learning in the output label space, although taking advantage of label correlations can boost the classification performance. In this paper, we propose a deep learning framework that incorporates flexible modules which can learn from both input and output spaces for multi-label image classification. For the input space learning, we devise a label-specific feature pooling method to refine convolutional features for obtaining features specific to each label. For the output space learning, we design a Two-Stream Graph Convolutional Network (TSGCN) to learn multi-label classifiers by mapping spatial object relationships and semantic label correlations. More specifically, we build object spatial graphs to characterize the spatial relationships among objects in an image, which supplements the label semantic graphs modelling the semantic label correlations. Experimental results on two popular benchmark datasets (i.e., Pascal VOC and MS-COCO) show that our proposed method achieves superior performance over the state-of-the-arts.
Keyword:
Feature extraction
Correlation
Task analysis
Semantics
Deep learning
Visualization
Benchmark testing
Multi-label image classification
label-specific feature
label correlations
graph convolutional network
deep learning
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
9.7
论文数:
4.5K
被引数:
2.4W
机构
引用论文
Multi-Label Metric Transfer Learning Jointly Considering Instance Space and Label Space Distribution Divergence联合考虑实例空间和标签空间分布散度的多标签度量迁移学习
IEEE ACCESS
IF3.6

