arrow
Return

prunAdag: an adaptive pruning-aware gradient method

delete2025-09-01
delete0
delete
OA
AI
M
Margherita Porcelli
G
Giovanni Seraghiti *
P
Philippe L. Toint
DOI:10.1007/s10589-025-00723-7delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
A pruning-aware adaptive gradient method is proposed which classifies the variables in two sets before updating them using different strategies. This technique extends the relevant/irrelevant approach of Ding et al. (Adv Neural Inf Process Syst 32, 2019) and Zimmer et al. (Mathematical optimization for machine learning: proceedings of the MATH+ thematic Einstein semester 2023, 2025) and allows a posteriori sparsification of the solution of model parameter fitting problems. The new method is proved to be convergent with a global rate of decrease of the averaged gradient's norm of the form \documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$\mathcal{O}(\log (k)/\sqrt{k+1})$$\end{document}. Numerical experiments on several applications show that it is competitive with existing pruning-aware Frank-Wolfe algorithms, see e.g. Zimmer et al. (Mathematical optimization for machine learning: proceedings of the MATH+ thematic Einstein semester 2023, 2025).
Keywords:
Model pruning
Adaptive first-order methods
Objective-function-free optimisation (OFFO)
Global convergence rate

Journal

C
Computational Optimization and Applications
IF:
2
Papers:
68
Citations:
3.5K

Organization

No organization information available
Cited Papers

Cited Papers

errShare
errSave
On the LambertW function
err1996-12-01
err0
PREAI
errR. M. Corless; G. H. Gonnet; D. E. G. Hare; D. J. Jeffrey; D. E. Knuth
errShare
errSave
researcher View more