arrow
返回

Improving Crowdsourcing-Based Image Classification Through Expanded Input Elicitation and Machine Learning

delete2022-06-29
delete4
delete
OA
AI
R
Romena Yasmin *
M
Md Mahmudulla Hassan
J
Joshua T. Grassel
H
Harika Bhogaraju
A
Adolfo R. Escobedo
O
Olac Fuentes
DOI:10.3389/frai.2022.848056delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
This work investigates how different forms of input elicitation obtained from crowdsourcing can be utilized to improve the quality of inferred labels for image classification tasks, where an image must be labeled as either positive or negative depending on the presence/absence of a specified object. Five types of input elicitation methods are tested: binary classification (positive or negative); the (x, y)-coordinate of the position participants believe a target object is located; level of confidence in binary response (on a scale from 0 to 100%); what participants believe the majority of the other participants' binary classification is; and participant's perceived difficulty level of the task (on a discrete scale). We design two crowdsourcing studies to test the performance of a variety of input elicitation methods and utilize data from over 300 participants. Various existing voting and machine learning (ML) methods are applied to make the best use of these inputs. In an effort to assess their performance on classification tasks of varying difficulty, a systematic synthetic image generation process is developed. Each generated image combines items from the MPEG-7 Core Experiment CE-Shape-1 Test Set into a single image using multiple parameters (e.g., density, transparency, etc.) and may or may not contain a target object. The difficulty of these images is validated by the performance of an automated image classification method. Experiment results suggest that more accurate results can be achieved with smaller training datasets when both the crowdsourced binary classification labels and the average of the self-reported confidence values in these labels are used as features for the ML classifiers. Moreover, when a relatively larger properly annotated dataset is available, in some cases augmenting these ML algorithms with the results (i.e., probability of outcome) from an automated classifier can achieve even higher performance than what can be obtained by using any one of the individual classifiers. Lastly, supplementary analysis of the collected data demonstrates that other performance metrics of interest, namely reduced false-negative rates, can be prioritized through special modifications of the proposed aggregation methods.
Keyword:
machine learning
input elicitations
crowdsourcing
human computation
image classification
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

F
Frontiers in Artificial Intelligence
IF:
4.7
论文数:
2.5K
被引数:
4.4K

机构

A
Arizona State University
学者数:
2.7W
论文数: 2.5W
被引数: 4.2W
A
arizona state university-tempe
学者数:
1.5W
论文数: 1.2W
被引数: 13
引用论文

引用论文

A solution to the single-question crowd wisdom problem
errNATURE
IF48.5
err2017-01-26
err190
PREAI
errPrelec, Drazen; Seung, H. Sebastian; Mccoy, John
err分享
err收藏
The Wisdom of Select Crowds
err2014-01-01
err186
PREAI
errMannes, Albert E.; Soll, Jack B.; Larrick, Richard P.
err分享
err收藏
Sugar, gravel, fish and flowers: Mesoscale cloud patterns in the trade winds
err2019-11-19
err96
errOAAI
errStevens, Bjorn; Bony, Sandrine; Brogniez, Helene; Hentgen, Laureline; Hohenegger, Cathy; Kiemle, Christoph; L'Ecuyer, Tristan S.; Naumann, Ann Kristin; Schulz, Hauke; Siebesma, Pier A.; Vial, Jessica; Winker, Dave M.; Zuidema, Paquita
err分享
err收藏
err分享
err收藏
学者 查看更多内容