arrow
Return

Detecting group concept drift from multiple data streams

delete2023-02-01
delete33
PRE
AI
H
Hang Yu
W
Weixu Liu
J
Jie Lü *
Y
Yimin Wen
X
Xiangfeng Luo
张广泉 (Guangquan Zhang)
DOI:10.1016/j.patcog.2022.109113delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Concept drift may lead to a sharp downturn in the performance of streaming in data-based algorithms, caused by unforeseeable changes in the underlying distribution of data. In this paper, we are mainly concerned with concept drift across multiple data streams, and in situations where the drift of each data stream cannot be detected in time, due to slight underlying distribution drifts. We call this group concept drift. When compared to the detection of concept drift for a single data stream, the challenges of detecting group concept drift arise from three aspects: first, the training data become more complex; second, the underlying distribution becomes more complex; and third, the correlations between data streams become more complex. To address these challenges, the key idea of our method is to construct a distribution free test statistic, free from any underlying distribution in multiple data streams. Then, for streaming data, we design an online learning algorithm to obtain this test statistic, thereby determining the concept drift caused by the hypothesis test. The experiment evaluations with both synthetic and realworld datasets prove that our method can accurately detect concept drift from multiple data streams.(c) 2022 Elsevier Ltd. All rights reserved.
Keywords:
Concept drift
Data streams
Online learning
Hypothesis test

Journal

Pattern Recognition cover
Pattern Recognition
IF:
7.6
Papers:
1.3W
Citations:
4.5W

Organization

U
university of technology sydney
Scholars:
1.6W
Papers: 2.0W
Citations: 25
G
Guilin University of Electronic Technology
Scholars:
7.4K
Papers: 5.2K
Citations: 5.4K
S
shanghai university
Scholars:
3.9W
Papers: 2.7W
Citations: 52
researcher View more organizations