arrow
Return

Controlled descent training

delete2024-01-16
delete1
delete
OA
AI
V
Viktor Andersson *
B
Balázs Varga
V
Vincent Szolnoky
A
Andreas Syrén
R
Rebecka Jörnsten
B
Balaźs Kulcsár *
DOI:10.1002/rnc.7194delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
In this work, a novel and model-based artificial neural network (ANN) training method is developed supported by optimal control theory. The method augments training labels in order to robustly guarantee training loss convergence and improve training convergence rate. Dynamic label augmentation is proposed within the framework of gradient descent training where the convergence of training loss is controlled. First, we capture the training behavior with the help of empirical Neural Tangent Kernels (NTK) and borrow tools from systems and control theory to analyze both the local and global training dynamics (e.g., stability, reachability). Second, we propose to dynamically alter the gradient descent training mechanism via fictitious labels as control inputs and an optimal state feedback policy. In this way, we enforce locally Script capital H2$$ {\mathscr{H}}_2 $$ optimal and convergent training behavior. The novel algorithm, Controlled Descent Training (CDT), guarantees local convergence. CDT unleashes new potentials in the analysis, interpretation, and design of ANN architectures. The applicability of the method is demonstrated on standard regression and classification problems.
Keywords:
convergent learning
gradient decent training
label augmentation
label selection
neural Tangent Kernel
optimal labels

Journal

International Journal of Robust and Nonlinear Control cover
International Journal of Robust and Nonlinear Control
IF:
3.2
Papers:
6.9K
Citations:
1.4W

Organization

C
chalmers university of technology
Scholars:
1.5W
Papers: 1.6W
Citations: 10