返回
Learning from multi-label data with interactivity constraints: An extensive experimental study
DOI:10.1016/j.eswa.2015.03.006.png)
摘要
En 中文
Interactive classification aims at introducing user preferences in the learning process to produce individualized outcomes more adapted to each user's behavior than the fully automatic approaches. The current interactive classification systems generally adopt a single-label classification paradigm that constrains items to span one label at a time and consequently limit the user's expressiveness while he/she interacts with data that are inherently multi-label. Moreover, the experimental evaluations are mainly subjective and closely depend on the targeted use cases and the interface characteristics. This paper presents the first extensive study of the impact of the interactivity constraints on the performances of a large set of twelve well-established multi-label learning methods. We restrict ourselves to the evaluation of the classifier predictive and time-computation performances while the number of training examples regularly increases and we focus on the beginning of the classification task where few examples are available. The classifier performances are evaluated with an experimental protocol independent of any implementation environment on a set of twelve multi-label benchmarks of various sizes from different domains. Our comparison shows that four classifiers can be distinguished for the prediction quality: RF-PCT (Random Forest of Predictive Clustering Trees, (Kocev, 2011)), EBR (Ensemble of Binary Relevance, (Read et al., 2011)), CLR (Calibrated Label Ranking, (Furnkranz et al., 2008)) and MLkNN (Multi-label kNIN, (Zhang and Zhou, 2007)) with an advantage for the first two ensemble classifiers. Moreover, only RF-PCT competes with the fastest classifiers and is therefore considered as the most promising classifier for an interactive multi-label learning system. (C) 2015 Elsevier Ltd. All rights reserved.
Keyword:
Interactive learning
Multi-label learning
Comparative study
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
7.5
论文数:
3.0W
被引数:
10.2W
机构
引用论文
Galaxy Zoo: morphologies derived from visual inspection of galaxies from the Sloan Digital Sky Survey银河动物园: 通过斯隆数字天空调查对星系进行目视检查得出的形态

