arrow
Return

CMAL: Cost-Effective Multi-Label Active Learning by Querying Subexamples

delete2022-05-01
delete12
delete
OA
AI
G
Guoxian Yu
X
Xia Chen
C
Carlotta Domeniconi
J
Jun Wang *
Z
Zhao Li
Z
Zili Zhang
X
Xiangliang Zhang
DOI:10.1109/TKDE.2020.3003899delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
Multi-label active learning (MAL) aims to learn an accurate multi-label classifier by selecting which examples (or example-label pairs) will be annotated and reducing query effort. MAL is a more complicated and expensive process than single-label active learning, due to one example can be associated with a set of non-exclusive labels and the annotator has to scrutinize the whole example and label space to provide correct annotations. Instead of scrutinizing the whole example for annotation, we may just examine some of its subexamples with respect to a label for annotation. In this way, we can not only save the annotation cost but also speedup the annotation process. Given this observation, we introduce CMAL, a two-stage Cost-effective MAL strategy (CMAL) by querying subexamples. CMAL first selects the most informative example-label pairs by leveraging uncertainty, label correlation and label space sparsity. Specifically, the uncertainty of a label to an example can be reduced if its correlated labels already annotated to the example, and its uncertainty can be reduced also if more examples annotated to this label. Next, CMAL greedily queries the most probable positive subexample-label pairs of the selected example-label pair. In addition, we propose rCMAL to account for the representative of examples to more reliably select example-label pairs in the first stage. Extensive experiments on multi-label datasets from diverse domains show that our proposed CMAL and rCMAL can better save the query cost than state-of-the-art MAL methods. The contribution of leveraging label correlation, label sparsity, and representative for saving cost is also confirmed.
Keywords:
Uncertainty
Correlation
Annotations
Measurement uncertainty
Training
Semantics
Multi-label active learning
multi-instance learning
label correlation
label sparsity
representative
uncertainty
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

IEEE Transactions on Knowledge and Data Engineering cover
IEEE Transactions on Knowledge and Data Engineering
IF:
10.4
Papers:
6.7K
Citations:
3.2W

Organization

A
alibaba group
Scholars:
1.1K
Papers: 789
Citations: 0
G
George Mason University
Scholars:
7.7K
Papers: 7.9K
Citations: 1.0W
S
southwest university - china
Scholars:
2.6W
Papers: 1.9W
Citations: 21
K
king abdullah university of science & technology
Scholars:
1.3W
Papers: 1.3W
Citations: 32
S
shandong university
Scholars:
9.3W
Papers: 6.4W
Citations: 94
researcher View more organizations