Return
Machine Learning From Crowds Using Candidate Set-Based Labeling
DOI:10.1109/MIS.2022.3205053.png)
Abstract
En 中文
Crowdsourcing is a popular and cheap alternative in machine learning for gathering information from a set of annotators. Learning from crowd-labeled data involves dealing with its inherent uncertainty and inconsistencies. In the classical framework, each annotator provides a single label per example, which fails to capture the complete knowledge of annotators. We propose candidate labeling, that is, to allow annotators to provide a set of candidate labels for each example and thus express their doubts. We propose an appropriate model for the annotators, and present two novel learning methods that deal with the two basic steps (label aggregation and model learning) sequentially or jointly. Our empirical study shows the advantage of candidate labeling and the proposed methods with respect to the classical framework.
Keywords:
Labeling
Reliability
Uncertainty
Intelligent systems
Task analysis
Computational modeling
Aggregates
Journal
IF:
6.1
Papers:
1.6K
Citations:
4.5K

