返回
Active learning through label error statistical methods
DOI:10.1016/j.knosys.2019.105140.png)
摘要
En 中文
Clustering-based active learning splits data into a number of blocks and queries the labels of the most critical instances. An active learner must decide how to choose these critical instances and how to split the blocks. In this paper, we present theoretical and practical statistical methods for analyzing the relationship between the label error and the neighbor radius, and design new split and selection strategies to handle these two issues. First, we define statistical functions for the label error based on a single instance and instance pairs. Second, we build practical statistical models, calculate empirical label errors, and guide the block splitting process. Third, using these practical models, we develop a center-and-edge instance selection strategy for choosing critical instances. Fourth, we design a new algorithm called active learning through label error statistical methods (ALSE). Learning experiments were performed with 20 datasets from various domains. The results of significance tests verify the effectiveness of ALSE and its superiority over state-of-the-art active learning algorithms. (C) 2019 Elsevier B.V. All rights reserved.
Keyword:
Active learning
Clustering
Label error statistical model
Probabilistic lipschitzness
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
K
IF:
7.6
论文数:
1.2W
被引数:
4.5W
机构
引用论文
Visual deficits and cognitive assessment of multiple sclerosis: confounder, correlate, or both?多发性硬化症的视觉缺陷和认知评估: 混杂,相关,或两者兼而有之?

