返回
Crowdsourced Data Management: A Survey
DOI:10.1109/TKDE.2016.2535242.png)
摘要
En 中文
Any important data management and analytics tasks cannot be completely addressed by automated processes. These tasks, such as entity resolution, sentiment analysis, and image recognition can be enhanced through the use of human cognitive ability. Crowdsouring platforms are an effective way to harness the capabilities of people (i.e., the crowd) to apply human computation for such tasks. Thus, crowdsourced data management has become an area of increasing interest in research and industry. We identify three important problems in crowdsourced data management. (1) Quality Control: Workers may return noisy or incorrect results so effective techniques are required to achieve high quality; (2) Cost Control: The crowd is not free, and cost control aims to reduce the monetary cost; (3) Latency Control: The human workers can be slow, particularly compared to automated computing time scales, so latency-control techniques are required. There has been significant work addressing these three factors for designing crowdsourced tasks, developing crowdsourced data manipulation operators, and optimizing plans consisting of multiple operators. In this paper, we survey and synthesize a wide spectrum of existing studies on crowdsourced data management. Based on this analysis we then outline key factors that need to be considered to improve crowdsourced data management.
Keyword:
Crowdsourcing
human computation
data management
quality control
cost control
latency control
期刊
IF:
10.4
论文数:
6.8K
被引数:
3.2W
机构
引用论文
Eggs sunny-side up: A new species of Olea, an unusual oophagous sea slug (Gastropoda: Heterobranchia: Sacoglossa), from the western Atlantic
Zootaxa
IF0
Design of Adaptive Fractional-Order Fixed-Time Sliding Mode Control for Robotic Manipulators
Entropy
IF0

