arrow
Return

Gradient preconditioned mini-batch SGD for ridge regression

delete2020-11-01
delete8
PRE
AI
Z
Zhuan Zhang
S
Shuisheng Zhou *
D
Dong Li
T
Ting Yang
DOI:10.1016/j.neucom.2020.06.092delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Data preconditioning technique, which reduces the condition number of the problem by a linear transformation of the data matrix, is typically used to accelerate the convergence of the first-order optimization methods for regularized loss minimization. One obvious limitation of the technique is exceedingly expensive of computational cost for the large-scale problems, especially an ocean of samples. In this paper, we have a gradient preconditioning trick and combine it with mini-batch SGD. The proposed gradient preconditioned mini-batch SGD algorithm boosts indeed the convergence with lower computational cost than that of the data preconditioning technique for ridge regression. Concretely, we use recent random projection and linear sketching methods to randomly low rank approximate the data matrix, then we can achieve a appropriate preconditioner through numerical linear algebra. Finally, we apply obtained preconditioner to the gradient to reduce computational cost. The experimental results on both synthetic data and real data sets validate the feasibility and effectiveness of our trick and algorithm. (C) 2020 Elsevier B.V. All rights reserved.
Keywords:
Mini-batch SGD
Regularized loss minimization
Ridge regression
Gradient preconditioning
Random projection
Linear sketching
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Neurocomputing cover
Neurocomputing
IF:
6.5
Papers:
2.5W
Citations:
6.5W

Organization

X
Xidian University
Scholars:
2.4W
Papers: 1.9W
Citations: 9.7K