arrow
Return

Classification utility aware data stream anonymization

delete2021-10-01
delete3
PRE
AI
O
Osman Abul *
DOI:10.1016/j.asoc.2021.107743delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Data streams are continuous, infinite and ordered sequences of data. In comparison to static dataset anonymization, data stream anonymization confront with a number of constraints and difficulties due to the dynamic nature of data flow. The literature already addressed the k-anonymization of data streams which contain quasi-identifier attributes. However, today most data streams contain sensitive and classification target attributes as well. This work's main motivation is to develop a k-anonymization method for data streams which additionally protects the sensitivity and enables effective classification models. The k-anonymization, as a result, is formulated as a weighted multi-objective optimization problem. There are three objectives with respective weights as user parameters. A clustering based k-anonymization algorithm is developed as the solution. An extensive experimental evaluation on three real datasets shows the effectiveness of our proposal in various configurations. Moreover, the experimental results also confirm that our proposal attains better classification accuracies in comparison to popular data stream anonymization techniques. (C) 2021 Elsevier B.V. All rights reserved.
Keywords:
Data streams
Data anonymization
Data privacy
Classification
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Applied Soft Computing cover
Applied Soft Computing
IF:
6.6
Papers:
1.4W
Citations:
4.8W

Organization

T
tobb ekonomi ve teknoloji university
Scholars:
957
Papers: 1.4K
Citations: 2