arrow
Return

Distributed Semi-Supervised Single-Index Model With Corruption

delete2025-11-04
delete0
PRE
AI
X
Xin Shen *
S
S Liu
J
Jiyuan Tu
DOI:10.1002/sta4.70115delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Modern dataset is often large in scale, necessitating the development of distributed training methods. In this paper, we focus on distributed learning for the single-index model (SIM). Particularly, we consider the presence of unlabelled covariate information and assume that an fraction of labels may be arbitrarily corrupted. To address the challenges of distributed training, we first propose a modified rank regression loss function that eliminates local bias in distributed training. To leverage the unlabelled covariates, we develop two distributed semi-supervised algorithms tailored for low-dimensional and high-dimensional settings, respectively. We theoretically demonstrate that the inclusion of unlabelled data accelerates distributed training. Notably, our method is inherently robust to a moderate fraction of label corruptions, regardless of how the corrupted labels are distributed across the worker machines. Simulation studies are provided to validate the effectiveness of our approach.
Keywords:
distributed learning
rank regression
robustness
semi-supervised learning
single-index model

Journal

S
STAT
IF:
0.8
Papers:
58
Citations:
655

Organization

S
shanghai jiao tong university
Scholars:
15.5W
Papers: 11.6W
Citations: 159
S
Shanghai University of Finance and Economics
Scholars:
2.0K
Papers: 2.5K
Citations: 4.0K