arrow
Return

Multi-label classification via incremental clustering on an evolving data stream

delete2019-11-01
delete38
delete
OA
AI
T
Tien Thanh Nguyen
T
Truong Dang
A
Anh Vu Luong
A
Alan Wee‐Chung Liew *
T
Tiancai Liang
J
John McCall
DOI:10.1016/j.patcog.2019.06.001delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
With the advancement of storage and processing technology, an enormous amount of data is collected on a daily basis in many applications. Nowadays, advanced data analytics have been used to mine the collected data for useful information and make predictions, contributing to the competitive advantages of companies. The increasing data volume, however, has posed many problems to classical batch learning systems, such as the need to retrain the model completely with the newly arrived samples or the impracticality of storing and accessing a large volume of data. This has prompted interest on incremental learning that operates on data streams. In this study, we develop an incremental online multi-label classification (OMLC) method based on a weighted clustering model. The model is made to adapt to the change of data via the decay mechanism in which each sample's weight dwindles away over time. The clustering model therefore always focuses more on newly arrived samples. In the classification process, only clusters whose weights are greater than a threshold (called mature clusters) are employed to assign labels for the samples. In our method, not only is the clustering model incrementally maintained with the revealed ground truth labels of the arrived samples, the number of predicted labels in a sample are also adjusted based on the Hoeffding inequality and the label cardinality. The experimental results show that our method is competitive compared to several well-known benchmark algorithms on six performance measures in both the stationary and the concept drift settings. (C) 2019 Elsevier Ltd. All rights reserved.
Keywords:
Multi-label classification
Incremental learning
Online learning
Clustering
Data stream
Concept drift
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Pattern Recognition cover
Pattern Recognition
IF:
7.6
Papers:
1.3W
Citations:
4.5W

Organization

R
Robert Gordon University
Scholars:
1.3K
Papers: 1.4K
Citations: 1.9K
G
Griffith University
Scholars:
1.5W
Papers: 1.6W
Citations: 2.5W
Cited Papers

Cited Papers

errShare
errSave
Beam search algorithms for multilabel learning
err2013-05-22
err71
errOAAI
errKumar, Abhishek; Vembu, Shankar; Menon, Aditya Krishna; Elkan, Charles
errShare
errSave
errShare
errSave
Scalable multi-output label prediction: From classifier chains to classifier trellises
err2015-06-01
err64
errOAAI
errRead, Jesse; Martino, Luca; Olmos, Pablo M.; Luengo, David
errShare
errSave
Multi-label classification via multi-target regression on data streams
err2016-12-30
err58
errOAAI
errOsojnik, Aljaz; Panov, Pance; Dzeroski, Saso
errShare
errSave
ML-KNN: A lazy learning approach to multi-label leaming
err2007-07-01
err2.8K
errOAAI
errZhang, Min-Ling; Zhou, Zhi-Hua
errShare
errSave
Perceived Racism as a Predictor of Paranoia Among African Americans
err2006-02-01
err0
PREAI
errDennis R. Combs; David L. Penn; Jeffrey Cassisi; Chris Michael; Terry Wood; Jill Wanner; Scott Adams
errShare
errSave
Decision trees for hierarchical multi-label classification
err2008-08-01
err497
errOAAI
errVens, Celine; Struyf, Jan; Schietgat, Leander; Dzeroski, Saso; Blockeel, Hendrik
errShare
errSave
researcher View more