arrow
Return

Learning to classify with missing and corrupted features

delete2009-07-16
delete83
delete
OA
AI
O
Ofer Dekel *
O
Ohad Shamir
L
Lin Xiao
DOI:10.1007/s10994-009-5124-8delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
A common assumption in supervised machine learning is that the training examples provided to the learning algorithm are statistically identical to the instances encountered later on, during the classification phase. This assumption is unrealistic in many real-world situations where machine learning techniques are used. We focus on the case where features of a binary classification problem, which were available during the training phase, are either deleted or become corrupted during the classification phase. We prepare for the worst by assuming that the subset of deleted and corrupted features is controlled by an adversary, and may vary from instance to instance. We design and analyze two novel learning algorithms that anticipate the actions of the adversary and account for them when training a classifier. Our first technique formulates the learning problem as a linear program. We discuss how the particular structure of this program can be exploited for computational efficiency and we prove statistical bounds on the risk of the resulting classifier. Our second technique addresses the robust learning problem by combining a modified version of the Perceptron algorithm with an online-to-batch conversion technique, and also comes with statistical generalization guarantees. We demonstrate the effectiveness of our approach with a set of experiments.
Keywords:
Adversarial environment
Binary classification
Deleted features

Journal

Machine Learning cover
Machine Learning
IF:
2.9
Papers:
2.7K
Citations:
3.4W

Organization

M
Microsoft
Scholars:
3.0K
Papers: 2.7K
Citations: 7
H
Hebrew University of Jerusalem
Scholars:
2.8W
Papers: 2.3W
Citations: 2.7W
Cited Papers

Cited Papers

Dynamical Cluster Approximation in Disordered Systems with Magnetic Impurities
err2004-12-15
err0
PREAI
errDaisuke Matsunaka; Hideaki Kasai; Wilson Agerico Diño; Hiroshi Nakanishi
errShare
errSave
2-D Material Molybdenum Disulfide Analyzed by XPS
err2014-07-09
err0
PREAI
errD. Ganta; S. Sinha; Richard T. Haasch
errShare
errSave
Gradient-based learning applied to document recognition
err1998-01-01
err3.8W
PREAI
errLecun, Y; Bottou, L; Bengio, Y; Haffner, P
errShare
errSave
Real-Time Analysis of a Modified State Observer for Sensorless Induction Motor Drive Used in Electric Vehicle Applications
err2017-07-25
err0
errOAAI
errMohan Krishna S.; Febin Daya J.L.; Sanjeevikumar Padmanaban; Lucian Mihet-Popa
errShare
errSave
researcher View more