arrow
Return

False selection rate control in mixture models

delete2025-09-01
delete0
delete
OA
AI
A
Ariane Marandon *
T
Tabea Rebafka
É
Étienne Roquain
N
Nataliya Sokolovska
DOI:10.1111/sjos.70017delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
The clustering task consists in partitioning elements of a sample into homogeneous groups. Most datasets contain individuals that are ambiguous and intrinsically difficult to attribute to one or another cluster. However, in practical applications, misclassifying individuals is potentially disastrous and should be avoided. To keep the misclassification rate small, one can decide to classify only a part of the sample. In the supervised setting, this approach is well known and referred to as classification with an abstention option. In this paper, the approach is revisited in an unsupervised mixture-model framework. The purpose is to develop a method that guarantees the false selection rate (FSR) does not exceed a predefined level . We propose a plug-in procedure and provide a theoretical analysis, quantifying the deviation of the FSR from the target with explicit remainder terms. Bootstrap versions of the procedure are shown to improve the performance in numerical experiments.
Keywords:
abstention option
bootstrap
clustering
false discovery rate
mixture models
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

S
Scandinavian Journal of Statistics
IF:
1
Papers:
52
Citations:
0

Organization

No organization information available